OpenAI rolls out upgraded prompt caching and lower cached input rates for GPT-6
rhiever · reddit · 2026-09-24
OpenAI announced upgraded prompt caching for GPT-6 along with reduced cached input rates, cutting costs for developers who repeatedly send long contexts such as system prompts and conversation history.
More from Models
- Why GPT-6's rumored recurrent-depth architecture could change inference economics — panic_in_the_galaxy · 2026-09-24
- Jev, a No-Text Model Claiming 200x Faster Decisions, Sparks 'Is a Well-Formatted Wrong Answer Still a Hallucination?' — jamesbrooksco · 2026-09-24
- Hands-on: Jev Beats Gemini Flash on Accuracy, Latency, and Cost — cantrell · 2026-09-24
- User cancels switch to Codex after Opus 5.5's 40% price cut wins him back — ayushtweetshere · 2026-09-24
- Codex completes steps Claude refuses twice in one day, dev reports — ChanceKelch · 2026-09-24
- Models Avoid Human Faces in 3D Benchmarks: FaceBench Still Too Hard — AymericRoucher · 2026-09-24