OpenAI launches Prompt Caching Dashboard and diagnostics API to track cache-hit rates

OpenAIDevs · x · 2026-09-23

OpenAI shipped a set of prompt-caching tuning tools for developers: explicit cache breakpoints to choose which prompt prefixes get reused, the ability to adjust reasoning effort and tool availability while preserving cached context, and prewarming of shared context for faster first responses. A new Prompt Caching Dashboard tracks cache-hit rates, and a diagnostics API identifies changes that broke reuse and estimates affected token counts.

Related event: OpenAI improves GPT-6 prompt caching with up to 90% input token savings(3 posts)→

Original post →

More from coding & agent

coding & agent channel →