OpenAI Developers (@OpenAIDevs) announced prompt-caching improvements for the GPT-6 API: higher default cache-hit rates so more input tokens get cached-input discounts of up to 90%, helping agents run faster and cheaper. OpenAI docs also point developers to the Prompt Caching Dashboard and miss Diagnostics.
Key Takeaways
- ✓GPT-6 API prompt caching now hits more often by default for agent workloads.
- ✓Cached input tokens discounted up to about 90%.
- ✓Dashboard monitors hit rates; Diagnostics help fix cache misses.

Discussion & Comments
0Sign in to join the discussion
Connect with AI developers to exchange benchmark insights.