Claude Fable 5 scores 210/210 on a bar exam benchmark for about $6
ctjlewis · x · 2026-07-22
Claude Fable 5 hits 210/210 on a bar exam benchmark
Matthew Stubenberg says Anthropic’s Claude Fable 5 became the first model to score perfectly on 210 MBE bar exam practice questions in his benchmark.
- The run cost about $6 in API tokens.
- It averaged roughly 10 seconds per question.
- Stubenberg notes the benchmark started in November 2022, when ChatGPT 3.5 was around 50% on the same set.
- His chart shows a steady climb toward perfect scores, with Fable 5 reaching 100%.
He frames it as evidence of how fast model performance on professional exams has improved in just over three years.
More from Models
- Hugging Face reportedly used open-weight GLM 5.2 after proprietary models failed — rasbt · 2026-07-22
- Google Exec Seeks Feedback on Gemini 3.6 Flash & 3.5 Flash-Lite Performance — patloeber · 2026-07-22
- Gemini 3.6 Flash is 2x faster and 18% cheaper, but independent tests say it is not smarter — etherd0t · 2026-07-22
- Critic says OpenAI incident coverage confuses bad reward functions with autonomy — ambaonadventure · 2026-07-22
- Google Launches Gemini 3.5 Flash Cyber Model for Security Teams — pushmeet · 2026-07-22
- Rumor says GPT-5.6 Sol could hit 750 tok/s after Cerebras upgrades — haider1 · 2026-07-22