Anthropic models now outpace DeepSeek-Flash on speed, but API TTFT looks suspicious
teortaxesTex · x · 2026-10-10
- The poster observes that, depending on how you measure, Anthropic's models are now faster than DeepSeek v4.1-flash — yet show curiously long TTFT on the API.
- He speculates this may be screening/rewriting followed by delayed token streaming, which could also help hide batching.
- The quoted post adds that v4.1-flash feels blazingly fast for building and iterating, slightly below GLM-5.3-flash on quality, and that switching back to other models feels painfully slow.
Related event: Anthropic Models Beat DeepSeek-Flash in Speed, Yet Show Odd API Latency(2 posts)→
More from Models
- Opus 5.5 Fast beats Sol 6.1 Ultrafast on price: 33% cheaper at the same ~300 tok/s speed — kimmonismus · 2026-10-10
- Developer slams Grok's 'nails on chalkboard' chat style and AI slop speech — max_paperclips · 2026-10-10
- Unannounced gpt-rosalind-discovery model spotted on OpenAI's API pricing page at $5/$25 per 1M tokens — testingcatalog · 2026-10-10
- Kimi K3.1 reportedly launching mid-October with long-horizon agent focus; K2.8 pricing raises concerns — teortaxesTex · 2026-10-10
- Kimi K3.1 not launching before mid-October, clarifies ChrisGPT — ChrisGPT · 2026-10-10
- Codex quality dip reported; GPT-61 Sol High said to be the better tier — vista8 · 2026-10-10