.05/AGI Hunt · AI 快讯 / AI 日报 · 实时更新的 AI 资讯站.16 per 1M… · AGI Hunt

DeepSeek V4 Flash listed at $0.05/$0.16 per 1M off-peak on OpenRouter, 4x below official pricing

michaelsoft__binbows · reddit · 2026-09-09

OpenRouter lists DeepSeek V4 Flash (0731) via openinference at $0.05/$0.16 per 1M tokens off-peak, versus DeepSeek's official $0.22/$0.66 — roughly 4x cheaper (official cached reads are $0.007 vs $0.013 on openinference).

The poster argues the price makes it hard to justify anything else: for tasks from chat summarization to complex reasoning, it undercuts smaller models (phi-4, qwen3.6 35B-A3B, qwen3.5-9B) on both cost and capability. Even though the author is building home GPU nodes to self-host models under 300B (targeting qwen3.8 27B and qwen3.8-flash-next for privacy), they concede this single API model is nearly impossible to beat without free electricity.

Original post →

More from Models

Models channel →