Home / Tools / Cache savings
Prompt cache savings calculator
Model how much you save when a share of input tokens hits a published cached-input rate. Satellite of the main TokenCalculator cost engine, not a second product.
OpenAI · input $5.00/1M · cached $2.50/1M
Share of input tokens billed at the cached rate (~6,400 tokens).
- No-cache request
- $0.04768
- With cache (80% hit)
- $0.03168
- Saved / request
- $0.016 (33.6%)
- Saved / month
- $480.00
Steady-state math: cached input at the published cache rate; remaining input at standard input; output unchanged. Write cost (when published) is a one-time prefix charge. Repay it with the per-request savings above. Rates verified 2026-08-12.
Next step
Paste a real prompt in the main calculator with the same model and cache slider to combine Exact/Approx counts with this savings math.
Open ChatGPT-4o Latest in cost calculator · Prompt caching guide · Batch pricing
Related tools
Related guides
Prompt caching explained · How prompt cost is calculated · How LLM API pricing works
Sources and references
Official documentation used for definitions, counting methods, or rate cards. Always confirm critical budgets on the provider page.