NewsOpenAI (official)Sep 22, 2026
Reducing latency and costs with prompt cache improvements in GPT-6
As improvements to GPT-6, higher prompt cache hit rate, new diagnostics, explicit breakpoints, and controls to reduce latency and cost will be added.
Why it mattersIn enterprise operations, responsiveness and cost management are essential. This update delivers immediate practical benefits.
Read the original →