GPT 6 Series Shows Frequent Lazy Work, While Sol 5.6 Spent Hours on QA
jdjohnson · x · 2026-10-01
- User jdjohnson reports that the GPT 6 series frequently produces lazy work with poor judgment on long tasks.
- By contrast, GPT Sol 5.6 was relentless: asked to turn an outline into a Keynote presentation matching his design system, it found the right skill and looped for hours on QA to keep the design consistent — wasteful but correct.
- Sol/Astra 6, however, invents its own design system and ships a sloppy PDF, which he calls "lazy and wrong."
More from Models
- Gemini element-naming race heats up: Neon and Argon taken, Krypton is next as Google rejoins the frontier — tkipf · 2026-10-01
- Rumor: Gemini 4 Argon will launch as Ultra subscriber exclusive at first — opmgyhx · 2026-10-01
- Researcher: Model's unprompted video-joke disclaimer is hard to explain without 'understanding' — technollama · 2026-10-01
- Talking math and physics with LLMs feels like working with an exam-acing savant — burny_tech · 2026-10-01
- One-line take: GPT-6.1 Sol is underrated, says AI commentator — mallow610 · 2026-10-01
- "Opus 5.5 is the cheapest model" — human hours saved beat token pricing — CamBrazy3 · 2026-10-01