MiniMax M3.1-Flash-Preview hits 150 tok/s and users find it distractingly fast
DanielLockyer · x · 2026-10-10
The author got access to MiniMax M3.1-Flash-Preview and found its 150 tok/s output feels instantaneous. He admits he isn't mentally prepared for fast models yet — he spends more time marveling at the speed than thinking about what to do next. A firsthand look at the usability shift preview-tier fast models bring.
More from Models
- Delip Rao calls out new Qwen3.5-9B-based model for benchmarking latency but not accuracy vs Jev — deliprao · 2026-10-10
- Gemini 4 Argon Launches: 77.9% on DeepSWE v1.1, Beating Claude Opus 5.5's 74.2% — dl_weekly · 2026-10-10
- Polymarket puts 28% odds on Anthropic pausing AI training this month — Polymarket · 2026-10-10
- Netlify livestream blind-tests new Anthropic, OpenAI and Mistral models on real tasks — thisiskp_ · 2026-10-10
- Google reportedly testing Gemini 4 "Carbon" internally; staff say coding "feels like Opus 5.5" — gaganghotra_ · 2026-10-10
- Max reasoning tier costs way more but scores worse on Terminal Bench 4, dev claims — weswinder · 2026-10-10