GPT-6's improved prompt caching cuts agent costs; leads AutomationBench at 88% less
OpenAIDevs · x · 2026-09-23
OpenAI detailed GPT-6's improved prompt caching: agents do less redundant processing, and more context stays reusable as reasoning effort and tool availability change, making responses faster and cheaper.
The same thread cites AutomationBench (measuring whether models can complete real business workflows across apps): GPT-6 Sol (xhigh) leads Fable 5.1 (max with Opus 5 fallback) with 88% lower reported cost per task.
Related event: OpenAI Launches GPT-6 Sol and Luna at Half the Price(48 posts)→
More from Infra
- vLLM ROCm maintainers finally get persistent AMD MI355X cluster after SemiAnalysis lobbying — AccBalanced · 2026-09-23
- NCCL 2.31.2 ships GPU-driven CFT, per-collective tuning and 0-SM collectives for Blackwell-scale training — SkyLi0n · 2026-09-23
- NCCL 2.31.2 ships GPU-driven CFT/RMA, 0-SM collectives, better multi-NIC for Blackwell-scale training — SkyLi0n · 2026-09-23
- vLLM ROCm lead maintainers only recently got persistent access to an MI355X cluster — AccBalanced · 2026-09-23
- South African Rights Groups Demand Moratorium on US Big Tech Data Centers — ChinasaTOkolo · 2026-09-23
- Delip Rao downgrades from $200/mo Google One Ultra to $50 Pro, leaning on local models — deliprao · 2026-09-23