Grow or Die

Free calculator · verified rate card

What does each user cost your AI product?

Use real calls and token counts to calculate cost per user, cost per call, and total model cost. Cache reads, cache writes, and retry attempts are priced separately.

Enter one period of model usage

All model rates and results are in USD.

Include failed or timed-out calls that are retried.
For OpenAI, subtract these from total prompt tokens before entering uncached input.
No synthetic token bill here.

Choose a model and enter your observed usage to calculate a real scenario. Empty fields stay empty.

Why cache and retries stay separate

Cached input is usually billed below standard input. Counting it at the full input rate overstates cost. This calculator keeps uncached input and cache reads as separate, non-overlapping buckets. Provider-reported cache writes are added at their own rate, and retries increase the number of paid attempts.