Ollama launches off-peak pricing: DeepSeek-V4 models half price nights and weekends
ollama · x · 2026-09-06
Ollama officially introduced off-peak token rates for its cloud inference: DeepSeek-V4-Flash and Pro are half price outside 12:00–18:00 UTC on weekdays (5–11am Pacific) and all day on weekends. More models will get off-peak pricing soon.
Regular per-million-token prices: deepseek-v4-flash at $0.22 input / $0.66 output, deepseek-v4-pro at $0.66 / $1.98, alongside gemma4, glm-5.x, kimi-k3, minimax and qwen3.5:397b. DeepSeek models are hosted in the US and Europe with zero data retention (ZDR). Subscription tiers include Free, Pro ($20/mo), Max ($100/mo) and Team ($500/mo).
Related event: Ollama Cloud Cuts DeepSeek-V4 Token Prices in Half Off-Peak(2 posts)→
More from Infra
- Ampere Public grants free access to ~1,000 interconnected chips for two-day research projects — charliermarsh · 2026-09-06
- Kimi and MiniMax to open Tmall stores selling token plans; Xiaomi releases table-data LDM — 创业邦 · 2026-09-06
- AMD exec: AI token processing could hit 120 quadrillion per month by 2030 — zephyr_z9 · 2026-09-06
- MiniMax-H3 Lip Sync on 8GB VRAM: Multishot Renders in 26 Minutes — big-boss_97 · 2026-09-06
- Oura's S-1 reveals an on-device AI stack: small models and edge compute, not cloud inference — eurie_kim · 2026-09-06
- Bosgame M5 Max With Ryzen AI Max Pro 495 and 192GB RAM Arrives October 2026 — Terminator857 · 2026-09-06