OpenAI Developers (@OpenAIDevs) announced prompt-caching improvements for the GPT-6 API: higher default cache-hit rates so more input tokens get cached-input discounts of up to 90%, helping agents run faster and cheaper. OpenAI docs also point developers to the Prompt Caching Dashboard and miss Diagnostics.

Key Takeaways

  • GPT-6 API prompt caching now hits more often by default for agent workloads.
  • Cached input tokens discounted up to about 90%.
  • Dashboard monitors hit rates; Diagnostics help fix cache misses.
Evaluating this AI coding model or solution?
Check live multi-benchmark rankings or compare plan costs & promo credits.
ADSponsored