Test: DeepSeek Harness achieves 99% cache hit rate with GLM and Kimi

sandyyevans · reddit · 2026-08-19

An experiment tested DeepSeek Harness (DSH) with GLM, Kimi, and Opus to verify if prompt caching works with non-DeepSeek models. Results showed GLM achieved 97% cache reuse in tool loops and 99.6% in subsequent turns, while Kimi hit 99% in both. This confirms DSH's append-style request pattern is compatible with other providers supporting prefix caching. Opus showed no cache activity in this test.

Original post →

More from Infra

Infra channel →