Prompt caching
Prompt Caching Savings Calculator
Compare uncached and cached input costs for repeated prompts, long system instructions, shared RAG context, and high-volume AI workflows.
Estimate prompt caching savings
Estimated monthly savings$0.00
Without caching$0.00
With caching$0.00
Savings rate0%
Cached input price$0.00
How caching changes the bill
Caching only helps when a meaningful part of the input is reused. The estimate comparesstandard input cost with cached reusable input + uncached dynamic input + output cost.
- Output tokens usually do not become cheaper through input caching.
- Provider cache rules can depend on model, minimum prefix length, and cache duration.
- Higher hit rates make caching more valuable.
FAQ
When does prompt caching save money?
Prompt caching helps when a meaningful portion of the input repeats across requests and the provider prices cached input below standard input.
Does prompt caching reduce output cost?
Usually no. Prompt caching lowers repeated input cost, while generated output tokens are still billed at the model's output price.
Related calculators
Prices were last checked on 2026-07-04. Verify provider cache rules because minimum lengths, TTLs, and write fees can differ.