OpenAI Boosts GPT-6 Prompt Caching, Input Tokens Now Up to 90% Cheaper
OpenAIDevs · x · 2026-09-23
OpenAI's official developer account announces improved prompt caching in the API for GPT-6, helping agents run faster and cost less. Higher cache-hit rates by default mean more input tokens benefit from cached-input discounts of up to 90%.
Related event: OpenAI improves GPT-6 prompt caching with up to 90% input token savings(3 posts)→
More from Models
- Unverified Early Impressions of Suspected Opus 5.5: Capable but 'Doesn't Feel Like an Opus' — repligate · 2026-09-23
- Artificial Analysis launches AA-Omniscience: all but 3 models hallucinate more than they answer right — geoffwolfe · 2026-09-23
- Anthropic launches Claude Opus 5.5: matches Fable 5.1 on most tasks at 40% lower cost — jyangballin · 2026-09-23
- AI researcher deletes poll on OpenAI math results, calling his framing misleading — ChrSzegedy · 2026-09-23
- Scale AI's Muse, built with Meta's Muse Spark 1.3, tops the App Store and PRBench — alexandr_wang · 2026-09-23
- Anthropic Opus 5.5 scores 93.3% on ARC-AGI-2 at 80% lower eval cost, ARC Prize confirms — burny_tech · 2026-09-23