GLM-5.3 Matches Claude Fable 5 on DeepSWE at 25% of the Cost
zainhas · x · 2026-08-23
Together Compute analyzed GLM-5.3 and Claude Fable 5 on the DeepSWE benchmark. GLM-5.3 matched Fable 5 in single-shot solve rate and pulled ahead with multiple attempts.
Key breakdown:
- Solve Rate: GLM-5.3 (69.0%) vs. Fable 5 (69.7%)
- Cost: GLM-5.3 ($3.99/task) vs. Fable 5 ($21/task)
- Tokens: GLM-5.3 (80k) vs. Fable 5 (114k)
- Turns: GLM-5.3 (124) vs. Fable 5 (85)
The analysis highlights that the economics become highly favorable when retries are cheap.
Related event: GLM-5.3 Matches Fable 5 on DeepSWE at a Quarter of the Cost(4 posts)→
More from Models
- AI Model Performance vs Cost: DeepSeek Flash and Gemini lead on value — bindureddy · 2026-08-23
- Why Closed AI Remains Silent After Qwen 3.8 27B Drop — My_Unbiased_Opinion · 2026-08-23
- Ex-Microsoft CTO Calls for OpenAI to Prioritize 'Pro' Model Config for Math — MParakhin · 2026-08-23
- AI Braille Understanding Test Shows Inconsistent Results Across Models — tristanbob · 2026-08-23
- Visualizing LLM Eval Stats: Improved Tables to Spot Bad CI Methods — IanArawjo · 2026-08-23
- DeepSeek V4 Pro model priced at around $0.19 amid quality praise — TheMoonMidas · 2026-08-23