Calculator
Check the number yourself.
If I sent you an estimate of your model spend, this is the page that produced it. Put your own figures in and the sum rewrites itself, line by line, so you can see exactly where the euros come from.
Nothing here is gated. No signup, no email, no credentials.
Your usage
Change anything here and every number on the right updates straight away.
Rates come from the same published price table Cost uses to bill your events.
The same calls and the same tokens, priced on a second model.
Start from a typical workload
These load ordinary figures for each shape of work. Change any of them below; the sum follows.
Every request your app makes to the model API, across all routes.
System prompt, retrieved context and the user message. Roughly 750 words is 1,000 tokens.
What the model writes back. Output usually costs about five times as much as input.
The share of your input tokens served from the prompt cache. If the model has no published cached rate, this changes nothing and the working says so.
Estimated model spend
€1,470per month
- Exact, in euro
- €1,469.70
- In dollars
- $1,597.50
- Per call
- $0.010650
- Per year
- €17,636
The working
Every step, in order. If we quoted you a figure by email, this is how we got it.
- 1Model: claude-sonnet-5 on Anthropic.
- 2Published rate: $2.00 per 1M input tokens, $10.00 per 1M output tokens, $0.20 per 1M cached input tokens.
- 3Usage: 150,000 calls a month, 3,000 input tokens and 600 output tokens per call.
- 4Cache: 25% of the input is a cache hit, so 750 tokens are billed at the cached rate and 2,250 at the full input rate.
- 5Input per call: 2,250 / 1,000,000 x $2.00 = $0.004500
- 6Cached input per call: 750 / 1,000,000 x $0.20 = $0.000150
- 7Output per call: 600 / 1,000,000 x $10.00 = $0.006000
- 8Cost per call: $0.010650
- 9Cost per month: 150,000 x $0.010650 = $1,597.50
- 10In euro at 0.92 USD to EUR: $1,597.50 x 0.92 = €1,469.70
- 11This counts API token cost only. It does not include cache-write charges, retries, embeddings, batch discounts or anything you spend outside the model API.
Converted at a standing rate of 0.92 USD to EUR. The dashboard uses the live daily rate instead, so your real invoice will differ by a percent or two.
What a cheaper model would cost
- claude-sonnet-5
- €1,469.70
- claude-haiku-4-5
- €734.85
€734.85a month less
€8,818 a year
Would quality hold?
On your traffic, nobody knows until it is tested. Most teams never switch, so they keep paying. Cost shadow-runs your real traffic against the cheaper model, scores it with an independent judge, and only recommends the switch after it passes.
That is a verified recommendation and a projected saving. It does not switch anything for you. Send a usage export or measure it on the free tier.
The working, compared
The same arithmetic on the second model, and the difference.
- 1Compared with: claude-haiku-4-5 on Anthropic, at exactly the same usage.
- 2Published rate: $1.00 per 1M input tokens, $5.00 per 1M output tokens, $0.10 per 1M cached input tokens.
- 3Cost per call on claude-haiku-4-5: $0.005325
- 4Cost per month: 150,000 x $0.005325 = $798.75, which is €734.85.
- 5Difference: €1,469.70 - €734.85 = €734.85 a month cheaper on claude-haiku-4-5.
- 6Over twelve months: €734.85 x 12 = €8,818.20.
- 7Same tokens and the same cache hit rate on both sides. A cheaper model may need longer prompts or more retries, and this does not model output quality.
Show the full sum for claude-haiku-4-5
- 1Model: claude-haiku-4-5 on Anthropic.
- 2Published rate: $1.00 per 1M input tokens, $5.00 per 1M output tokens, $0.10 per 1M cached input tokens.
- 3Usage: 150,000 calls a month, 3,000 input tokens and 600 output tokens per call.
- 4Cache: 25% of the input is a cache hit, so 750 tokens are billed at the cached rate and 2,250 at the full input rate.
- 5Input per call: 2,250 / 1,000,000 x $1.00 = $0.002250
- 6Cached input per call: 750 / 1,000,000 x $0.10 = $0.000075
- 7Output per call: 600 / 1,000,000 x $5.00 = $0.003000
- 8Cost per call: $0.005325
- 9Cost per month: 150,000 x $0.005325 = $798.75
- 10In euro at 0.92 USD to EUR: $798.75 x 0.92 = €734.85
- 11This counts API token cost only. It does not include cache-write charges, retries, embeddings, batch discounts or anything you spend outside the model API.
This model has an announced price rise
From 2026-09-01, the same usage costs €2,205 a month, which is €734.85 more than today.
- Introductory pricing ends 2026-08-31. Input and output both rise 50%.
- New rate from 2026-09-01: $3.00 per 1M input tokens (was $2.00), $15.00 per 1M output tokens (was $10.00).
- Same usage at the new rate: $2,396.25 a month, which is €2,204.55.
- That is €734.85 a month more than today, an increase of 50%.
Source: anthropic.com, checked 2026-08-01. If that page now says something else, believe it over this one.
Want the full breakdown by route?
I can send you a written report that takes this estimate apart: which model tiers you could drop to, what prompt caching would actually save you, and where the same spend usually hides in a codebase like yours. Leaving an email is entirely optional. The calculator above is complete without it and nothing on this page is hidden behind a form.
What this estimate does not include
- Retries and failed calls. Most teams are surprised by these, and they are usually worth a few percent.
- Cache writes. Populating a prompt cache costs more than a plain input token, so a heavily cached workload pays a small write charge this page ignores.
- Embeddings, fine-tuning, batch discounts and anything billed outside the model API.
- The spread between your assumed token counts and your real ones. That gap is the whole reason Cost meters the live calls instead of guessing, and it is usually bigger than people expect.
If you would rather not guess at token counts, upload a usage export on the free analysis page and I will work from your real numbers. Or wrap one client with the SDK and see it measured, on the free tier.