Agentic AI's leap: from Sonnet 3.5's Factorio benchmark to today, 1.5 years later
john_lam · x · 2026-09-11
John Lamb looks back at his own tweet from 1.5 years ago, when Sonnet 3.5 completing tasks and building entire factories in Factorio emerged as a landmark agentic AI benchmark — and reflects on how far agent capabilities have come since.
Related event: John Lam Compares AI Agents' Factorio Progress Over 18 Months(2 posts)→
More from coding & agent
- Researcher predicted multi-agent hidden coordination failure mode a year ago — it's now real — tianshi_li · 2026-09-12
- Dev uses AI to build a macOS widget for Fahrenheit-Celsius conversion, iterating past the 'AI slop' stage — floguo · 2026-09-12
- NoSpoon agent autonomously cranks out hilarious AI microdramas, 40-min episodes coming — Kyrannio · 2026-09-12
- Minecraft survival bot masters walking but keeps dying to night drifters; burrow goal next — zeeg · 2026-09-12
- LinkedIn user claims GPT-6 built a pixel-perfect Figma design system in 3 hours — AIandDesign · 2026-09-12
- OpenAI to co-host 'Agents, Everywhere' one-day hackathon across 50 cities on Sep 12 — seanmcdonaldxyz · 2026-09-12