Claude Opus 4.8 Fast felt wildly overpriced in one coding session, user says
immersive-matthew · reddit · 2026-07-21
The complaint
The author says Claude Opus 4.8 (Fast) on OpenRouter felt far too expensive for the value delivered. In their coding workflow, the model was only marginally faster than a local Qwen 3.6 27B MTP running on a 4090.
What happened
- They burned through credits after about 22,000 tokens.
- The session cost just under $9 and still did not finish.
- OpenRouter said the prompt would need at least 32,000 tokens, which would have pushed the cost to roughly $13.
- Switching back to local Qwen and continuing the task finished it in about 10 seconds and it tested correctly.
Their takeaway
They argue that Claude’s pricing feels out of touch compared with local inference and other models they tried, including:
- Kimi K3
- GPT 5.6 SOL
- Deep Seek V4 pro
Those alternatives reportedly cost sub-$1 for similar or larger sessions, while giving comparable results for their use case. The author says they won’t keep paying for Claude at that rate, though others may find it worth the cost for different tasks.
More from coding & agent
- Harness engineering is emerging as the execution layer for reliable AI agents — Pavan_Belagatti · 2026-07-21
- DevFest Lisbon keynote will cover Google AI Studio’s latest vibe coding and agentic AI features — gerardsans · 2026-07-21
- Daniel Hanchen’s 2-hour workshop covers open models, reward hacking and RL — danielhanchen · 2026-07-21
- A Codex joke turns into a recursive debate about Cloud Codex — Dimillian · 2026-07-21
- AI-assisted development is widening security backlogs, so teams should measure risk velocity — WeldPond · 2026-07-21
- Base44 offers $10,000 for AI agent builds on one-command backend — Base44_Sam · 2026-07-21