Qwen 27B Coding Run Consumes 919M Input Tokens, Caching Cuts Costs

TheZachMueller · x · 2026-08-25

Developer stats for DeepSWE runs on Qwen2.5 27B show 919M input tokens and 8M output tokens across 10,580 agent turns. While the input volume looks high, the author notes that KV Cache allows reuse across turns, making it far cheaper than processing fresh tokens. In comparison, the most inefficient config (Claude Code high thinking) took nearly 5 days on an RTX 6000.

Original post →

More from coding & agent

coding & agent channel →