HF daily papers: DeepSeek KV cache at 890 bytes/token, auto-research loop cuts agent tokens ~45-49%

ThomasAger · reddit · 2026-09-19

The author highlights the top 3 papers on HF Daily Papers, all relevant to local LLMs and agent harness optimization:

An unusually dense trio covering KV compression, automated harness optimization, and harness design ablations.

Original post →

More from Research

Research channel →