OpenAI disclosed improvements to prompt caching for GPT-6, introducing higher cache hit rates, explicit breakpoints, new diagnostics, and additional controls designed to reduce latency and costs1. The update expands developer tooling around cache behavior, giving API users more granular visibility into how cached prefixes are matched and reused.
OpenAI Upgrades GPT-6 Prompt Caching with Breakpoints, Diagnostics
OpenAI disclosed GPT-6 prompt caching improvements including higher cache hit rates, explicit breakpoints, new diagnostics, and controls to reduce…