User Tests New Model: Strong at Fermi Estimates and Strict Poetry Rewriting
jd_pressman · x · 2026-08-22
A user shared their test experience with a new model. Key findings include:
- Fermi Estimates: The model recalled and processed an astonishing number of facts and figures to complete the estimate.
- Poetry Rewriting: When asked to rewrite Blake's 'Jerusalem' in anapestic pentameter, the model thought for a long time and largely succeeded.
- Self-Awareness: The model knew the user to a lesser degree compared to other frontier models.
Related event: Community tests new model: impressive but brand bias persists(2 posts)→
More from Models
- GLM 5.3, Fable 5, and GPT-5.6 Sol show opposite results on Terminal-Bench 3 vs DeepSWE — zainhas · 2026-08-22
- Claude interrogates you to guess your vibe; Grok just reads your tweets — repligate · 2026-08-22
- Relying solely on benchmarks and consensus fails to capture true model capabilities — nptacek · 2026-08-22
- Opus 5 allocates skills to coding, philosophy, and understanding human intent — davidad · 2026-08-22
- Fable 5 excels at postdoc-level math, reversing Anthropic's historical underperformance — davidad · 2026-08-22
- Frontier model capabilities are jagged; custom evals for specific use cases are essential — nptacek · 2026-08-22