From the source
Lead story
Top stories
Models & availability
Latest
Lead story
Top stories
Models & availability
Latest
From the source
Improved caching and new monitoring tools for GPT-6 agents
OpenAI improved prompt caching for GPT-6 with higher cache hit rates by default within a 30-minute window.
The company introduced a Prompt Caching Dashboard to monitor cache performance, a diagnostics tool to identify cache misses, and new controls including explicit cache breakpoints, reasoning effort adjustment without breaking cache, and cache prewarming to help developers reduce latency and costs.
From the source
With the GPT‑6 family, we launched an improved prompt caching system that delivers higher cache hit rates by default. We now give cache discounts for eligible shared prefixes reused within a 30-minute window. We're also introducing new tools to help developers monitor cache performance, diagnose misses, and choose how much of a prompt to cache.
openai.com