Better prompt caching for GPT-6
GPT-6 improves prompt caching with higher hit rates, explicit breakpoints, and new latency/cost controls.
OpenAI announced prompt caching improvements for GPT-6, including higher cache hit rates, new diagnostics, explicit breakpoints, and controls to reduce latency and cost. As a power-player release referencing GPT-6, it signals both a next-gen model and infrastructure-level efficiency gains relevant to developers building on OpenAI.