Only OpenAI Mastered Reasoning, Yet DeepSeek Tops Benchmark with Medium Effort
teortaxesTex · x · 2026-08-03
AI KOL @teortaxesTex argues that DeepSeek's discovery of RL for reasoning in r1 is independent of OpenAI's o1 technology, asserting that only OpenAI has truly mastered intrinsic 'reasoning effort'.
Meanwhile, the quoted tweet reveals the latest VulcanBench results: DeepSeek took the top spot, surprisingly achieving this with medium effort rather than high effort. Grok 4.5 Medium ranked second, while ChatGPT fell out of the top five.
More from Models
- Speechify's Simba Tops Voice Blind Test Leaderboard at 1/10th of Competitors' Price — PrajwalTomar_ · 2026-08-03
- Minimax H3 still missing, Reddit user jokes 'Where is it?' — Opposite-Answer5456 · 2026-08-03
- Minimax open-source model imminent? Time up but not yet released — Pitiful_Archer_4381 · 2026-08-03
- Developer experiments with fine-tuning models to generate WIP animations — johnowhitaker · 2026-08-03
- Why Are Frontier LLMs Getting Worse? A Deep Dive into Modern Training Pipelines — MarcJSchmidt · 2026-08-03
- Gary Marcus: OpenAI's Astra is Vastly Oversold, Math Breakthroughs Don't Mean AGI — Gary Marcus · 2026-08-03