Switching models constantly invalidates prompt cache and doubles costs

Teknium · x · 2026-08-25

Teknium warns that constantly switching models in a session invalidates the prompt cache on the new model, forcing you to repay the full input token price for all context. This is a fundamental inference principle; avoid doing this unless the models are free.

Original post →

More from coding & agent

coding & agent channel →