Claude Opus 4.8 Fast felt wildly overpriced in one coding session, user says
immersive-matthew · reddit · 2026-07-21
The complaint
The author says Claude Opus 4.8 (Fast) on OpenRouter felt far too expensive for the value delivered. In their coding workflow, the model was only marginally faster than a local Qwen 3.6 27B MTP running on a 4090.
What happened
- They burned through credits after about 22,000 tokens.
- The session cost just under $9 and still did not finish.
- OpenRouter said the prompt would need at least 32,000 tokens, which would have pushed the cost to roughly $13.
- Switching back to local Qwen and continuing the task finished it in about 10 seconds and it tested correctly.
Their takeaway
They argue that Claude’s pricing feels out of touch compared with local inference and other models they tried, including:
- Kimi K3
- GPT 5.6 SOL
- Deep Seek V4 pro
Those alternatives reportedly cost sub-$1 for similar or larger sessions, while giving comparable results for their use case. The author says they won’t keep paying for Claude at that rate, though others may find it worth the cost for different tasks.
More from coding & agent
- GPT-6 Astra beats Factorio with enemies in 44 in-game hours at ~$4,500 API cost — liminal_bardo · 2026-09-11
- 105 hidden bugs, 2 repos: DeepSeek V4.1 Flash fixes 24 at $1.80 vs Opus 5's 27 at $51.33 — ChartsJournalX · 2026-09-11
- Investment Analyst Asks How to Build a Claude-Based Diligence Agent Stack — Careless_Tie2286 · 2026-09-11
- Treating agents like 50 First Dates: a 3-layer context system so every conversation doesn't start from zero — evielync · 2026-09-11
- Running the Firefox MCP on Android via Termux, ngrok, and mcp-proxy — Nervous-Strain7544 · 2026-09-11
- SmolVM open-sources persistent computer infrastructure for agents that outlive chat sessions — aniketmaurya · 2026-09-11