DeepSeek compressed KV cache 54x in nine months, analysts say

Analysts note DeepSeek has cut per-token KV cache 54x in nine months to 890 bytes, boosting token efficiency toward OpenAI-style levels, with long-context benchmarks awaited.

2026-09-10 ~ 2026-09-10 · 4 related posts