AI Development
OpenAI improves prompt caching for GPT-6 with higher hit rates and new tools
Editorial Analysis
OpenAI launched an improved prompt caching system for the GPT-6 family that delivers higher cache hit rates by default and provides discounts of up to 90% on cached input tokens for eligible shared prefixes reused within a 30-minute window. New developer tools include a Prompt Caching Dashboard, diagnostics for cache misses, explicit cache breakpoints, the ability to adjust reasoning effort without breaking cache, and prewarming support, aimed at reducing latency and cost for persistent multi-turn agents.
At a Glance
Date
September 22, 2026
Importance
Medium
3/5
Category
inference
Axis of Change
Cost Reduction
Organizations
Models Affected
Sources
- 01 openai.com