DeepSeek API Summary Billing Raises Questions
teortaxesTex · x · 2026-07-19
This post raises a specific API billing question: if the platform summarizes long CoT for display but charges based on the full token output, does the displayed "short summary" actually reduce real costs?
The author seeks clarification on the DeepSeek official API platform's actual behavior, arguing that if billing is based on the original output length despite a short summary, users are still paying for the full reasoning output. The post also casually notes that credible V4 GA-related outputs on Bilibili visually pale in comparison to Fable and Kimi.
More from Models
- Grok 4.5 is now free inside Cursor, the popular AI coding IDE — mark_k · 2026-07-21
- GPT often converges on the same near-miss ideas in math problems — yacineMTB · 2026-07-21
- Eno Reyes says model distillation is basically unstoppable — LangChain · 2026-07-21
- Sakana says multiple diffusion models plus MCTS beat test-time scaling on coding and math — SakanaAILabs · 2026-07-21
- OpenAI hackathon project stalls as Codex struggles on voice, while Claude spots the issue — ColleenMBrady · 2026-07-21
- Kimi K3 lands exactly on China’s 2-year AI capability trend line — peterwildeford · 2026-07-21