Haiku 5.5 Hits 67.1 on BuildingBench, Beating Sonnet 5 at 1/11 the Cost
ZhitingHu · x · 2026-10-10
EnactraAI benchmarked two new models on BuildingBench, a construction-focused coding benchmark:
- Haiku 5.5 (max) jumped from its predecessor's 26.7 to 67.1, beating Sonnet 5 by 13 points at roughly 1/11 the cost: $1.34 vs. $15.03 per building.
- Mistral Large 4 (max) debuted at 52.0 using the Claude Code harness, at $33.21 per building — cost-efficiency has room to improve.
- Haiku 5.5 narrowly misses the Pareto frontier: SWE-2 scores almost identically and is still free during its promotion.
More from Models
- Datology launches Curation Studio, stress-tested with hundreds of on-demand B300s on Modal — josh_wills · 2026-10-10
- Leak claims major model drops next week, dismissing slowdown talk — iruletheworldmo · 2026-10-10
- Haiku flubs trading basics: can't tell bids from offers, user reports — arthurcolle · 2026-10-10
- Step 5 Preview builds a working dashboard with zero code changes, demo shows — StepFun_ai · 2026-10-10
- HiDream-O1-Video-1.0 debuts #6 on image-to-video leaderboard at $5.80/min — ArtificialAnlys · 2026-10-10
- Blogger Speculates Anthropic Had Opus 5.5-Level Intelligence Internally Six Months Ago — yihui_indie · 2026-10-10