OpenAI details free ChatGPT upgrades: major factual errors down 65%, sycophancy down 83%
michpokrass · x · 2026-09-10
An OpenAI staffer lays out how the company improves the default ChatGPT experience for its 1B+ weekly users, most of whom use it for free, aiming to keep the default experience close to the capability frontier.
Key gains over the past 6 months:
- Responses with major factual errors down 65%, down 72% in high-stakes categories like finance
- GPT-5.6 Sol (instant) and GPT-5.6 Luna (medium) beat o3 at high reasoning effort on GPQA diamond while being 30%+ faster to last token
- Extreme sycophancy down 80%, overall sycophancy down 83%
- 83% fewer hallucination-flagged answers on a high-stakes medical eval, translating to hundreds of millions of better health conversations monthly
- New-user cohorts pursue 20-30% more use cases at least three times weekly
Free users now get unlimited text chats, higher reasoning effort, scheduled task Automations, and better personalization via "dreaming" memory. The author also outlines a "personal AGI" vision: progressive, model-assisted onboarding so ordinary users can naturally reach the highest-utility AI applications, since hard-to-learn tech fails to reach broad audiences.
Related event: OpenAI Says Major Factual Errors Down 65%, Sycophancy Down 83%(2 posts)→
More from Companies & People
- Safety researcher joins OpenAI's Safety and Security Committee, stresses externally verifiable oversight — JacquesThibs · 2026-09-10
- Waymo co-CEO Dmitri Dolgov, at Google's self-driving project since 2009, to speak at SPC — adityaag · 2026-09-10
- Midjourney opens annual community call to rank next 12 months of priorities — midjourney · 2026-09-10
- Alexandr Wang's Muse app hits #3 on the App Store — alexandr_wang · 2026-09-10
- Warner Music forces Suno to kill its original AI model, lobbies governments on AI training — Neurogence · 2026-09-10
- Paul Christiano joins OpenAI board's Safety and Security Committee — ethanCaballero · 2026-09-10