GPT-6 prompt caching can be tuned with breakpoints and prewarming
OpenAI Developers says applications can select reusable prompt prefixes with explicit cache breakpoints, change reasoning effort and tool availability while preserving cached context, and prewarm shared context so responses start sooner.