Ollama cloud launches off-peak token pricing: DeepSeek-V4 at half price
ollama · x · 2026-09-06
Ollama's cloud service introduced off-peak token rates: DeepSeek-V4-Flash and DeepSeek-V4-Pro are now half price outside weekdays 12:00–18:00 UTC (5–11am Pacific) and all day on weekends, with plans to extend the scheme to more models. Ollama notes its hosted DeepSeek models run in the US and Europe with ZDR (zero data retention) and fast performance.
Related event: Ollama Cloud Cuts DeepSeek-V4 Token Prices in Half Off-Peak(2 posts)→
More from Infra
- How much VRAM do you actually need for local document AI like paperless-ai? — pjdonovan · 2026-09-06
- 96-plasmid yeast transformation in 3.25 hours with open-source Pylabrobot — nlarusstone · 2026-09-06
- Dev Claims Fix for NVIDIA's Intentional P2P Nerf on Consumer GPUs — QuixiAI · 2026-09-06
- Fine-tuning LLMs in the browser: WebGPU training PoC built on llama.cpp — ngxson · 2026-09-06
- DGX Spark driver fixes reclaim 64GB of 'missing' GPU memory and boost page faults 49x — QuixiAI · 2026-09-06
- Ollama cloud full price list: $0.015 to $15 per million tokens across models — paw_lean · 2026-09-06