Ways to buy 6
| Way to buy | Input | Output | In / out per 1M | Cached | Vs cheapest |
|---|---|---|---|---|---|
Default region | $0.188 | $0.653 | $0.188 / $0.653Cheapest | n/a | Cheapest |
base, fp8 | $0.2 | $0.8 | $0.2 / $0.8+15% vs cheapest | n/a | +15% vs cheapest |
fp8 | $0.27 | $0.85 | $0.27 / $0.85+37% vs cheapest | n/a | +37% vs cheapest |
fp8 · cached $0.17 | $0.35 | $1 | $0.35 / $1+69% vs cheapest | $0.17 | +69% vs cheapest |
us-east5 | $0.35 | $1.15 | $0.35 / $1.15+81% vs cheapest | n/a | +81% vs cheapest |
From the platform pricing page | $0.371 | $1.48 | $0.371 / $1.48+114% vs cheapest | n/a | +114% vs cheapest |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $31.80 direct.
Helicone$31.80+0%
Kong AI Gateway$31.80+0%
LiteLLM$31.80+0%
Vercel AI Gateway$31.80+0%
Cloudflare AI Gateway$33.39+5%
Requesty$33.39+5%
Eden AI$33.55+5.5%
OpenRouter$33.55+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Meta · 1M | $1.25 | $4.25 | $1.25 / $4.25 | $2 | 1M | Lab list price | · |
Meta · 1M | $0.1 | $0.2 | $0.1 / $0.2 | $0.125 | 1M | Lab list price | · |
Meta · 1M | $0.1 | $0.2 | $0.1 / $0.2 | $0.125 | 1M | Lab list price | · |
Meta · 131K | $0.3 | $1.1 | $0.3 / $1.1 | $0.5 | 131K | on Phala | · |
DeepSeek · 1MCheapest: $0.1 on Relace | $0.15 | $0.6 | $0.15 / $0.6 | $0.262 | 1M | $0.1 on Relace | · |
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.281 on Google Gemini API (flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Gemini API (flex) | · |
Cohere · 192K | $0.3 | $1.5 | $0.3 / $1.5 | $0.6 | 192K | Lab list price | · |
About Llama 4 Maverick
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Takes text, image in and returns text. First listed Sep 23, 2026. Up to 16K output tokens. Blended price $0.304 per 1M tokens at 3 input to 1 output. Model id llama-4-maverick.
Price history
No list-price history: this model is sold by hosts rather than its lab.
Questions
How much does Llama 4 Maverick cost?
$0.188 per million input tokens and $0.653 per million output tokens.
What is the context window of Llama 4 Maverick?
1M tokens, with up to 16K tokens of output.
Where is Llama 4 Maverick cheapest?
Of 6 routes tracked, DigitalOcean is cheapest at $0.188 input and $0.653 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.