LLM API cost calculator
Price a workload across 168 models from 14 creators, each at its cheapest provider.
1M tokens is about 752K words or 2,506 pages.
Estimated cost by model
For 798K input and 200K output tokens, the cheapest priced model is Llama 3.1 8B at $0.026 via Novita; the most expensive is o1-pro at $239.70. Cheapest first.
How the cost is calculated
Formula
Input tokens times the model's input rate, plus output tokens times its output rate, both per 1M tokens. Words and pages convert to tokens at the ratio printed under the fields.
Prices
Each model is priced at the cheapest provider listing it, with input and output rates from that same provider. Rates are base-tier, on-demand prices in USD per 1M tokens; cached-input, batch, long-context and reasoning-token billing are not modelled. A model missing either rate is not ranked.
Order, sorting and filtering
Cheapest first for the workload entered, and the list opens on the 25 cheapest; "Show all" reveals the rest, and so does any search, filter or sort. Click a column header to sort by it, again to reverse it, and a third time to go back to cost. The search, the creator chips and "Open weights" combine: a model has to match all three. The bar beside each cost is that cost against the most expensive model in the list.
Transparency and funding
- Affiliates: Affiliate links are marked. We may earn a commission if you click them, but commissions never affect the order.
| Model | Cheapest provider | Context | Input /1M | Output /1M | Cost |
|---|---|---|---|---|---|
|
|
Novita | 131K | $0.02 | $0.05 | $0.026 |
|
|
Novita | 8K | $0.03 | $0.03 | $0.030 |
|
|
Alibaba Cloud | 1M | $0.03 | $0.13 | $0.050 |
|
|
Geodd | 131K | $0.03 | $0.14 | $0.052 |
|
|
Novita | 128K | $0.05 | $0.10 | $0.060 |
|
|
Cohere | 128K | $0.04 | $0.15 | $0.062 |
|
|
Cloudflare | 128K | $0.03 | $0.20 | $0.064 |
|
|
Novita | 128K | $0.04 | $0.17 | $0.066 |
|
|
Geodd | 131K | $0.04 | $0.18 | $0.068 |
|
|
Geodd | 262K | $0.05 | $0.20 | $0.080 |
|
|
Replicate | 8K | $0.05 | $0.25 | $0.090 |
|
|
Lyceum | 262K | $0.06 | $0.25 | $0.098 |
|
|
Mistral | 256K | $0.10 | $0.10 | $0.100 |
|
|
Cloudflare | 128K | $0.05 | $0.34 | $0.11 |
|
|
Cloudflare | 33K | $0.05 | $0.34 | $0.11 |
|
|
DigitalOcean | 1M | $0.08 | $0.25 | $0.11 |
|
|
Z.AI | 1M | $0.08 | $0.25 | $0.11 |
|
|
OpenAI | 400K | $0.05 | $0.40 | $0.12 |
|
|
Lyceum | 128K | $0.10 | $0.30 | $0.14 |
|
|
Cloudflare | 256K | $0.10 | $0.30 | $0.14 |
|
|
Lyceum | 33K | $0.10 | $0.30 | $0.14 |
|
|
EmpirioLabs AI | 262K | $0.07 | $0.42 | $0.14 |
|
|
EmpirioLabs AI | 262K | $0.06 | $0.46 | $0.14 |
|
|
Mistral | 256K | $0.15 | $0.15 | $0.15 |
|
|
Google Cloud | 1M | $0.10 | $0.40 | $0.16 |
No models matching your filters.
Showing 25 of 168 models
Heads up: Base-tier, on-demand rates per 1M tokens; cached, batch and long-context tiers excluded. A provider may serve a shorter context or a quantized build than the creator's release. Rates are taken from the cheapest provider listing each model; the bar is each cost against the most expensive model in the list. Verify before provisioning. More on how we price.
What common jobs cost
Each job is a stated assumption about tokens, priced at every model's cheapest provider. Pick one to load it into the calculator.
| Job | Tokens | Cheapest model | Cheapest frontier model |
|---|---|---|---|
| Summarising a 200-page book 80K in · 798 out | 81K | $0.002 Llama 3.1 8B | $0.16 Qwen3.8-Max |
| Translating a 5,000-word article 7K in · 7K out | 13K | under $0.001 DeepSeek OCR-2 | $0.053 Qwen3.8-Max |
| Classifying 100,000 support tickets 20M in · 100K out | 20M | $0.41 Llama 3.1 8B | $40.60 Qwen3.8-Max |
| Running a chatbot through 10,000 conversations 66M in · 13M out | 80M | $2.00 Llama 3.1 8B | $212.80 Qwen3.8-Max |
| Extracting fields from 1,000 invoices 798K in · 200K out | 998K | $0.026 Llama 3.1 8B | $2.80 Qwen3.8-Max |
| Summarising 50 hours of meeting transcripts 598K in · 20K out | 618K | $0.013 Llama 3.1 8B | $1.32 Qwen3.8-Max |
| Reviewing a 10,000-line pull request 80K in · 1K out | 81K | $0.002 Llama 3.1 8B | $0.17 Qwen3.8-Max |
Common questions
How is the cost calculated?
Input tokens times the model's input rate, plus output tokens times its output rate, both per 1M tokens. The rate is the cheapest provider listing the model, with input and output taken from that one provider. Cached-input discounts, batch pricing, long-context surcharges and reasoning-token billing are not modelled, because the rates we track do not include them; a workload that uses them will cost less, or more, than the figure here.
How many words is 1 million tokens?
About 751,880 English words, or roughly 2,506 book pages at 300 words a page. This page converts at 1.33 tokens per word, OpenAI's rule of thumb that a token is about three quarters of a word; Gemini quotes 60 to 80 words per 100 tokens. Claude 4.7 and later tokenize about 30% denser than earlier Claude models, and text in languages other than English can take 1.5x to 3x the tokens, so treat the conversion as an estimate and count with the provider's tokenizer before relying on it.
How many tokens are in a 200-page book?
About 80K tokens: 200 pages at 300 words is 60,000 words, times 1.33 tokens a word. A 300-page book is about 120K tokens and a typical page about 399. Dense technical text, tables and code run higher.
What does 1 million tokens cost?
Per 1M input tokens, the models on this page run from $0.020 (Llama 3.1 8B) to $150.00 (o1-pro); the median is $0.44. Output tokens cost more, typically 3x to 8x the input rate, so a job that writes as much as it reads costs several times a job that only reads.
Which provider's price is used for each model?
The cheapest one listing it. Open-weight models are served by several providers at different rates, sometimes severalfold apart; closed models resold through partners are usually at the creator's list price. Every model name links to its page, which lists every provider we track for it.
How current are these prices?
Rates were last updated on September 15, 2026 and are reviewed against each provider's published pricing page by hand. The full table, with every provider, is on the LLM API pricing page.
Rates last updated .
← All LLM prices