PayAPIKey

Prompt Cache Savings Calculator

Compare uncached and cached input token costs and estimate monthly prompt caching savings.

prompt cache calculatorcached token cost calculatorAI API cache savings

Prompt Cache Savings Calculator

Compare normal and cached input pricing for repeated prompt content.

Without cache
$1,250
Normal input pricing
With cache
$483
70% hit rate
Monthly savings
$768
61.4%
Effective rate
$0.97/1M
Blended cached input rate
Cached token volume350M
Uncached token volume150M
Break-even hit rate1.8%
Caching saves $768 per month at the current hit rate. Verify cache-write, storage, minimum-prefix, and expiration rules for your provider.

Frequently asked questions

Practical answers for applying this calculator to a production API billing or usage plan.

How does prompt caching reduce API cost?

Eligible repeated input tokens are billed at a lower cached-input rate. Savings depend on the provider's cache rules, the share of reusable prompt content, and the achieved cache hit rate.

Does a high cache hit rate guarantee savings?

Only when cached reads are cheaper than normal input and cache write or storage charges do not outweigh the discount. Model those extra charges separately when they apply.

Which tokens should count as cacheable?

Use stable system prompts, tool definitions, long reference documents, and repeated conversation prefixes that meet the provider's caching requirements.