Ways to buy 3
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
The lab's own API · cached $0.1 | $0.4 | $1.6 | $0.4 / $1.6List price | $0.1 | List price |
Default region · cached $0.1 | $0.4 | $1.6 | $0.4 / $1.6Same as list | $0.1 | Same as list |
swedencentral · cached $0.11 | $0.44 | $1.76 | $0.44 / $1.76+10% vs list | $0.11 | +10% vs list |
USD per 1M tokens; batch $0.2 in, $0.8 out on the lab's API. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $72.00 direct.
Helicone$72.00+0%
Kong AI Gateway$72.00+0%
LiteLLM$72.00+0%
Vercel AI Gateway$72.00+0%
Cloudflare AI Gateway$75.60+5%
Requesty$75.60+5%
Eden AI$75.96+5.5%
OpenRouter$75.96+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $2 on OpenAI (flex) | $2 | $10 | $2 / $10 | $4 | 1.05M | $2 on OpenAI (flex) | · |
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Z.ai · 1.31M | $0.563 | $2.5 | $0.563 / $2.5 | $1.05 | 1.31M | on DeepInfra (fp4) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Z.ai · 205K | $0.43 | $1.75 | $0.43 / $1.75 | $0.76 | 205K | on Venice (fp4) | · |
About GPT-4.1 Mini
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost.
Takes image, text, files in and returns text. First listed Apr 15, 2025. Up to 33K output tokens. Blended price $0.7 per 1M tokens at 3 input to 1 output. Model id gpt-4-1-mini.
Price history
List price since Apr 15, 2025, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.
Questions
How much does GPT-4.1 Mini cost?
$0.4 per million input tokens and $1.6 per million output tokens, with cached input at $0.1, or $0.2 and $0.8 through the Batch API.
What is the context window of GPT-4.1 Mini?
1.05M tokens, with up to 33K tokens of output.
Where is GPT-4.1 Mini cheapest?
Of 3 routes tracked, Azure AI Foundry is cheapest at $0.4 input and $1.6 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.