Gemini 3.7 Flash and GPT 5.6 Luna ranked best for automation tasks
burkov · x · 2026-08-30
After comparing cost, quality, predictability, and consistency across various tasks, the author identifies Gemini 3.7 Flash and GPT 5.6 Luna as the top models for automation. No open-weight models at a similar price point were found to be comparable, raising concerns about a potential future void if these models are retired.
More from Models
- Qwen 350K Context Tested on M5 Max: Performance and Quality — Artistic_Okra7288 · 2026-08-30
- GLM-5.3 vs. Flash: 17x Price Difference and Usage Strategy — togethercompute · 2026-08-30
- Grok $300/Month Subscriber Reports Hitting Only 10% of Weekly Usage Cap — AaronBergman18 · 2026-08-30
- Benchmarking models by hand takes forever, but I care about the data and model welfare — cephaloform · 2026-08-30
- Model Suspected of Text RL, Creates Own Marketing — isidentical · 2026-08-30
- I built a guide to the “Best LLMs for Coding” using 11 benchmark boards — DataLearnerAI · 2026-08-30