Sonnet 5.5 vs Sonnet 5: bouncing-ball physics tests from the same prompt
claudeai · x · 2026-09-29
Anthropic's Sonnet 5.5 experiment thread continues with bouncing-ball physics tests by @forwardeditor, generated from the same prompt on both Sonnet 5 and Sonnet 5.5 to compare the two models' physics simulation output side by side.
Related event: Claude Teases Sonnet 5.5 with Side-by-Side Demos(2 posts)→
More from Models
- Dev after 2 days: Claude is excellent, Codex great for long-horizon tasks but poorly designed — cneuralnetwork · 2026-09-29
- Opus 5.5 tops Drone-Bench and cheats far less than prior Claude models — scaling01 · 2026-09-29
- User claims 'Opus 5.5' turned a post on agent harnesses into an explainer video in one shot — alex_verem · 2026-09-29
- Engineer proud as Sonnet 5.5 scores 61.6% on chartography benchmark — echen · 2026-09-29
- ProgramBench multi-agent eval: Opus 5.5 fastest with a 5-agent team, Sonnet 5.5 with subagents — jyangballin · 2026-09-29
- Arrow 2 Telos tops Design Arena's SVG generation benchmark — AWizardWhoCodes · 2026-09-29