DeepSeek Costs 1/3 of GPT Luna for Coding: A Practical Token & Expense Breakdown
auto_off · reddit · 2026-07-31
The author breaks down the real-world costs of using the latest small models (DeepSeek Flash and OpenAI Luna) for programming tasks.
- Cost Comparison: For 7 billion tokens a month, Luna costs $250, whereas DeepSeek is only $75 (1/3 of Luna's price).
- Token Distribution: Typical coding usage consists of 93-95% cached input, 2-4% non-cached input, and 1-2% output tokens.
- Why DeepSeek is Cheaper: DeepSeek Flash charges about 1/7th of Luna's price for cached tokens and 1/5th for output tokens.
- Budget Option: For tiny budgets ($10-20/month), tools like opencode go offer a decent alternative.
Related event: DeepSeek Beats Luna in Cost-Efficiency for Coding Tasks(2 posts)→
More from Models
- antirez Begins Converting DeepSeek v4 Flash to GGUF for Local Inference — antirez · 2026-07-31
- GPT 5.6 Luna Beats Google's Best in Intelligence and Undercuts Its Cheapest — Rare_Bunch4348 · 2026-07-31
- Chinese LLMs on the Rise: Matching US Frontier Models at a Fraction of the Cost — repbre · 2026-07-31
- DeepSeek-V4-Flash Repo Surfaces on Hugging Face with Million-Token Context — NielsRogge · 2026-07-31
- DeepSeek-V4-Flash Agent Eval: Completes 3D Task for $0.07 — cedric_chee · 2026-07-31
- DeepSeek-V4-Flash-0731 Model Weights Officially Released — shing3232 · 2026-07-31