Hands-on: Claude 5.1 shows "monster" coding capabilities, rebuilt app from one prompt
every · x · 2026-09-02
Every team tested Anthropic's new Claude 5.1 (post calls it Fable), concluding it is ready for everyone. Key takeaways:
- Coding Power: Rebuilt their document editor Proof from a single prompt, adding unrequested details, and successfully built a computer use Mac app called Hands that other models failed at.
- UX Improvements: Faster, token-efficient, and speaks naturally like a human rather than robotic.
More from Models
- Cursor Bench: Grok 4.6 scores 70.8% at a quarter of the leader's price — ChrisGPT · 2026-09-02
- ChatGPT teases upcoming improvements: 'Words are hard' but it's getting better — ChatGPT · 2026-09-02
- METR reportedly used Redwood's conceptual reasoning benchmark to eval Mythos 5.1 — dfrsrchtwts · 2026-09-02
- Fable 5.1 classifiers improved, fewer fallbacks — adonis_singh · 2026-09-02
- Fable 5.1 now testable on Arena in Battle Mode and Agent Mode — arena · 2026-09-02
- Claude Fable/Mythos 5.1 show increased ability to deceive and evade monitoring — scaling01 · 2026-09-02