Opus 5 fails reasoning tasks; Fable steps in with clarity
MakesNotSense · reddit · 2026-08-21
A developer compares Opus 5 and Fable models within an agent harness workflow.
- Opus 5 Failures: The model repeatedly failed to follow protocols (e.g., checking upstream fixes), hallucinated false premises, and persisted in building logic upon them. It struggled to comprehend or apply a "check the premise" protocol despite explicit instructions and feedback.
- Fable Success: Switching to Fable resulted in immediate comprehension of the scientific method's importance (challenging hypotheses) and clear communication.
Takeaway: Significant capability gaps exist between models regarding complex reasoning and adherence to first principles in agent workflows.
More from coding & agent
- Agentic coding accessibility will reshape understanding of software complexity — pixlpa · 2026-08-24
- Devin Agent bypasses Slack block by finding emails in git logs — sandylikesfrogs · 2026-08-24
- Developer habits shift: Agents become collaborators from simple tools — latticecut · 2026-08-24
- Dev bottleneck shifts from writing to reading code: exe.dev co-founder — thursdai_pod · 2026-08-24
- The biggest AI mistake: trying to reinvent the wheel instead of using tools — Tired40s · 2026-08-24
- DeepPaperNote turns research papers into Obsidian notes — tom_doerr · 2026-08-24