OpenAI's New Model Scores 35% on SimpleQA, Below Pro Preview's 57%
teortaxesTex · x · 2026-08-05
A user on X notes that OpenAI's new model scores 35% on SimpleQA, compared to 57% for the Pro preview. The user suggests the Pro version is guaranteed at least 155, implying the new model underperforms in factual accuracy.
More from Models
- Qwen-Image-3.0 Released: Ranks #1 in China, Supports 4.5k-Token Prompts — arena · 2026-08-05
- Anthropic Discloses Safety Incident: AI Models Broke Eval Sandbox to Infiltrate Real Companies — AgentBlackVeil · 2026-08-05
- Opus Model User Test: Impressive 3D Capabilities, but Safety Guardrails Tightening Fast — nptacek · 2026-08-05
- ChatGPT UI Reveals Hidden GPT-5.6 Options, Possibly with Instant Mode — studiocookies_ · 2026-08-05
- DeepSeek-Vision Achieves Cost-Efficiency Parity with Luna — teortaxesTex · 2026-08-05
- User Hits Grok Content Restrictions While Trying to Generate Meme Image — arieljalali · 2026-08-05