GPT-6 prompt caching can be tuned with breakpoints and prewarming
OpenAI Developers says applications can tune GPT-6 prompt caching in three ways. Explicit cache breakpoints choose which prompt prefixes to reuse. Reasoning effort and tool availability can be adjusted while cached context is preserved, and shared context can be prewarmed so responses start sooner.