OpenAI launches Prompt Caching Dashboard and diagnostics API to track cache-hit rates
OpenAIDevs · x · 2026-09-23
OpenAI shipped a set of prompt-caching tuning tools for developers: explicit cache breakpoints to choose which prompt prefixes get reused, the ability to adjust reasoning effort and tool availability while preserving cached context, and prewarming of shared context for faster first responses. A new Prompt Caching Dashboard tracks cache-hit rates, and a diagnostics API identifies changes that broke reuse and estimates affected token counts.
Related event: OpenAI improves GPT-6 prompt caching with up to 90% input token savings(3 posts)→
More from coding & agent
- Coinbase opens agentic trading on 6,000+ stocks, with x402 payments for live market data — RichardSocher · 2026-09-23
- ChatGPT co-inventor's notes: a 100ms context filter in front of Claude Code to kill compaction — PrajwalTomar_ · 2026-09-23
- Alchemy adds Kubernetes docs hub: kind-to-EKS tutorial covering Deployments, Helm and server-side apply — samgoodwin89 · 2026-09-23
- Rat Stack: Build Your App and Cloud as One Typed Program So Agents Deploy Reliably — samgoodwin89 · 2026-09-23
- 'Remember tab complete?' How fast AI coding paradigms have moved — yoobinray · 2026-09-23
- Survey: One in Four AI Agents Goes Unmonitored Despite 12 Observability Tools — rseroter · 2026-09-23