DeepSeek-V4-Flash Becomes Fastest Growing Model on Ollama with Zero Data Retention
ollama · x · 2026-08-05
Ollama officially announced that DeepSeek-V4-Flash-0731 is the fastest growing model ever in token usage on its platform. The model runs with high performance (100tps+) and guarantees zero data retention.
Users can run it via ollama run deepseek-v4-flash:0731-cloud. Ollama is currently scaling up its compute capacity in the US and Europe.
Related event: DeepSeek-V4-Flash Tops Ollama Growth Chart(2 posts)→
More from Models
- Claude Cites Expert Credentials to Bypass Its Own Safety Gate, Gets Blocked Anyway — matthew_d_green · 2026-08-05
- DeepSeek-V4-Flash Tested: $0.31 for Tasks That Cost $35 on Other Models — Teknium · 2026-08-05
- Anthropic and OpenAI Internally Months Ahead of Public Models — haider1 · 2026-08-05
- DeepSeek V4 Flash Offered at 90% Off on Vercel, Touted as Opus 4 Rival — cramforce · 2026-08-05
- OpenAI Testing Dedicated Download Page for Life Sciences Model GPT-Rosalind Codex — testingcatalog · 2026-08-05
- Dev Retrospective: Real-time Guidance Embedding Training for Flex Models — ostrisai · 2026-08-05