Free calculator · verified rate card
What does each user cost your AI product?
Use real calls and token counts to calculate cost per user, cost per call, and total model cost. Cache reads, cache writes, and retry attempts are priced separately.
Enter one period of model usage
All model rates and results are in USD.
How to read it
Why cache and retries stay separate
Cached input is usually billed below standard input. Counting it at the full input rate overstates cost. This calculator keeps uncached input and cache reads as separate, non-overlapping buckets. Provider-reported cache writes are added at their own rate, and retries increase the number of paid attempts.