OpenAI Upgrades GPT-6 Prompt Caching With Higher Hit Rates and New Diagnostics
OpenAI News · rss · 2026-09-23
OpenAI announced improved prompt caching for GPT-6, featuring higher cache hit rates, new diagnostics, explicit breakpoints, and additional controls that reduce latency and costs for API workloads.
More from Infra
- $500 of Dell OptiPlexes become a diskless netboot lab where AI agents can't brick the hardware — colinmcnamara · 2026-09-23
- AMD MI355X beats NVIDIA B200 by 2.25x at scale with ~40% lower cost per GPU hour: Signal65 — ryanshrout · 2026-09-23
- 2.3B MoE hybrid Mamba model matches Llama-3.2-3B with <1% of its pretraining FLOPs — tri_dao · 2026-09-23
- ByteDance Seed fixes full-pipeline FP8 RL instability with Calibrated Clipping across 8B-32B models — ByteDance-Seed · 2026-09-23
- DigitalOcean Managed Agents Enters Public Preview: Pause-When-Idle Cloud Claude Code and Codex — _AustinCalvert_ · 2026-09-23
- ARK analyst: AI is the most powerful joule in history, converting energy to GDP ~10x better than humans — DMaguireARK · 2026-09-23