Background
Text size
Audio
Speaker also sits top-left ↖
A+ Larger textis on
The independent AI pricing desk

AI pricing, measured
for your workload.

Compare model costs. Track price changes.

22 providers Checked nightly Rates up to 500× apart

Compare prices
3 /22
Claude Sonnet 5
$0.00186
under a penny

Play with AI pricing.Tap the wheel to lock in a model.

At 1,000,000 emails a day, that's $69.16 to $5,320 depending on the model.

or keep exploring below

spin · pick a task · tap centre
Try another provider and watch the chart change Try another provider

See how a conversation compounds 10 replies re-read the growing history
Live from your picks above
Every reply re-reads the entire conversation
200 words of context with 150-word replies. Change the wheel provider or size above and this redraws.
Anthropic Claude Sonnet 5 Email · 200 words
Current mid-tier rates: $2 input / $10 output per 1M tokens. Choose another provider in the wheel to compare.
R1$0.00293
R2$0.00332
R3$0.00372
R4$0.00412
R5$0.00452
R6$0.00492
R7$0.00532
R8$0.00572
R9$0.00612
R10$0.00652
10 replies total$0.04722
Context re-read (grows)
Output written (fixed)

What does this cost at scale? Use the same selected workload at monthly volume
📈 Scale your selection
Take the exact 10-reply conversation above, then choose how many times it runs in a month.
10-reply conversations per month3,000
Your selected model
Anthropic · Claude Sonnet 5 $141.65 / month
3,000 email conversations, each with 10 replies

Across all 66 standard-rate model tiers, that same monthly workload ranges from $1.06 to $708.22 . Same volume, different model.

Explore the market DeepInfra leads the cheapest mid-tier list this week
📊 What leads this week Cheapest mid-tier model for a typical job · rank change vs 7 days ago
  1. 1DeepInfra$0.000878
  2. 2Cohere$0.00140
  3. 3Fireworks AI$0.00140
  4. 4Groq$0.00140
  5. 5OpenRouter$0.00219
Compare all 22 providers →
What does it actually cost to run an AI API call?

It depends on the model and how much text goes in and out. TokenScale shows real examples. Writing the whole of The Hobbit costs about $0.32 on Gemini Flash-Lite 3.5 at list rates as of 9 Sep 2026, and about $0.0051 on the cheapest tracked tier. The same job costs far more on a frontier model. Pick a content size and a provider to see the live figure.

Why does the output price matter more than the input price?

Most providers charge more for the tokens a model generates than for the tokens you send. Output often costs three to five times more than input. A long answer to a short question can cost more than it looks. TokenScale splits every price into input and output so the gap is visible.

How much do AI prices differ between providers?

A lot. Across the 22 providers TokenScale tracks, input rates span roughly 500×, from $0.02 to $10 per million tokens as of 9 Sep 2026, and even models aimed at similar work are routinely more than thirty times apart. Choosing the right provider for a task can cut the bill by an order of magnitude. The comparison table sorts every provider cheapest first.

What is the cheapest way to run a large AI workload?

Pick a low-cost provider and then check that model’s own batch or cached-input terms if your workload can use them. Discounts and eligibility vary by model and provider. TokenScale compares standard on-demand observations, marks carried values, and uses exact model-specific Batch prices where public terms are established; it never applies one universal discount. Open-weight models on inference hosts are often among the cheapest.

Is TokenScale free?

Yes. TokenScale is free, with no sign-up and no account. It runs entirely in your browser. Nothing you type is sent anywhere or stored.

How often is the pricing updated?

Every night. TokenScale checks provider pricing sources and records what the check could reproduce. A successful observation is verified; otherwise the last known value is carried forward and marked for that date.

4.1 ★★★★☆ AI Critics