On August 20, 2026, OpenAI released a Prompt Caching dashboard on its API platform1. It surfaces cache hit rate over time, cache reads per write, and the breakdown of cache-read, cache-write and uncached tokens1.
Metrics can be filtered by model and by service tier1. OpenAI describes the purpose as understanding caching efficiency and identifying opportunities to improve it1. Prompt caching bears directly on an API bill, yet how often it was actually hitting has been hard to track. The company also cut the price of its flagship GPT-5.6 Sol on August 21, so costs are moving on both the unit-price and the efficiency side.
Sources
- Changelog - OpenAI API - OpenAI official API documentation (entry for August 20, 2026)