LLM API pricing
253 models, 1,143 routes, 22 ways to buy, per 1M tokensUpdated 4h ago
30 of 253 models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Anthropic · 1M | $10 | $50 | $10 / $50 | $20 | 1M | Lab list price | · |
OpenAI · 1.05MCheapest: $10 on OpenAI (flex) | $10 | $50 | $10 / $50 | $20 | 1.05M | $10 on OpenAI (flex) | · |
Anthropic · 1M | $5 | $25 | $5 / $25 | $10 | 1M | Lab list price | · |
Anthropic · 1M | $10 | $50 | $10 / $50 | $20 | 1M | Lab list price | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
Anthropic · 1M | $5 | $25 | $5 / $25 | $10 | 1M | Lab list price | · |
Anthropic · 1M | $5 | $25 | $5 / $25 | $10 | 1M | Lab list price | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Anthropic · 1M | $4 | $20 | $4 / $20 | $8 | 1M | Lab list price | · |
OpenAI · 1.05MCheapest: $5.63 on OpenAI (flex) | $5 | $30 | $5 / $30 | $11.3 | 1.05M | $5.63 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $33.8 on OpenAI (flex) | $30 | $180 | $30 / $180 | $67.5 | 1.05M | $33.8 on OpenAI (flex) | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Google · 1MCheapest: $2.25 on Google Vertex AI (global-flex) | $2 | $12 | $2 / $12 | $4.5 | 1M | $2.25 on Google Vertex AI (global-flex) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Qwen (Alibaba) · 1M | $2 | $6 | $2 / $6 | $3 | 1M | Lab list price | · |
Moonshot AI · 1M | $1.4 | $10.8 | $1.4 / $10.8 | $3.74 | 1M | on InferenceNet (fp4) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Z.ai · 1.31M | $0.563 | $2.5 | $0.563 / $2.5 | $1.05 | 1.31M | on DeepInfra (fp4) | · |
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Anthropic · 200K | $5 | $25 | $5 / $25 | $10 | 200K | Lab list price | · |
OpenAI · 1.05MCheapest: $2.25 on OpenAI (flex) | $2 | $12 | $2 / $12 | $4.5 | 1.05M | $2.25 on OpenAI (flex) | · |
Anthropic · 1M | $3 | $15 | $3 / $15 | $6 | 1M | Lab list price | · |
OpenAI · 1.05MCheapest: $33.8 on OpenAI (flex) | $30 | $180 | $30 / $180 | $67.5 | 1.05M | $33.8 on OpenAI (flex) | · |
xAI · 500K | $1.6 | $4.8 | $1.6 / $4.8 | $2.4 | 500K | Lab list price | · |
Qwen (Alibaba) · 1M | $2 | $6 | $2 / $6 | $3 | 1M | Lab list price | · |
Anthropic · 1M | $2 | $10 | $2 / $10 | $4 | 1M | Lab list price | · |
Anthropic · 1M | $3 | $15 | $3 / $15 | $6 | 1M | Lab list price | · |
Z.ai · 205K | $0.43 | $1.75 | $0.43 / $1.75 | $0.76 | 205K | on Venice (fp4) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
DeepSeek · 1MCheapest: $0.1 on Relace | $0.15 | $0.6 | $0.15 / $0.6 | $0.262 | 1M | $0.1 on Relace | · |
Latest price changes
All changes| Seen | Model | Old price | New price | Change |
|---|---|---|---|---|
| Sep 24, 2026 | Input $0.66Output $1.98Cached input $0.022 | Input $1.32Output $3.96Cached input $0.044 | +100% | |
| Sep 24, 2026 | Input $0.15Output $0.6Cached input $0.003 | Input $0.3Output $1.2Cached input $0.006 | +100% | |
| Aug 29, 2026 | Input $0.375Output $1.88Cached input $0.037 | Input $0.75Output $3.75Cached input $0.075 | +100% | |
| Aug 24, 2026 | Input $2.5Output $15Cached input $0.25 | Input $2Output $10Cached input $0.2 | -20% | |
| Aug 24, 2026 | Input $2.5Output $15Cached input $0.25 | Input $2Output $10Cached input $0.2 | -20% | |
| Aug 21, 2026 | Input $5Output $30Cached input $0.5 | Input $2.5Output $15Cached input $0.25 | -50% |
Ways to buy model access
Compare allQuestions
What is the cheapest LLM API?
Sort the models page by blended price, or read the lower-left corner of the price map. The cheapest capable models are usually the small "flash", "mini" or "lite" tiers and open-weight models served by hosts; the cheapest route for an open model is often a host rather than the lab.
How are LLM API prices quoted?
In US dollars per million tokens, with separate prices for input (the prompt) and output (what the model writes). Output usually costs 3 to 8 times input. Cached input and batch requests are cheaper where a provider offers them.
What does "blended" price mean here?
One number for ranking: three parts input to one part output, the usual mix for chat and retrieval work. Your own mix may differ; the calculator uses your actual token counts.
Where do the prices come from?
OpenRouter's public model catalog and its per-provider endpoint list, read every morning at 05:00 UTC, which carries each provider's list price. Each provider page links to the official pricing page. Price history before tracking began comes from Internet Archive snapshots of the same catalog.
Is GPT or Claude cheaper?
It depends on the tier. Compare them side by side on the compare page, which shows input, output, cached and batch prices and every route (the lab, Azure, Bedrock, Vertex).
Do AI gateways charge extra?
Some do. OpenRouter adds a fee when you buy credits, Vercel AI Gateway says it adds no markup, and self-hosted proxies like LiteLLM add nothing but your own hosting. The gateways page lists each one with its source.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.