API pricing per million tokens
- Input
- $0.1
- per 1M, list
- Output
- $0.4
- per 1M tokens
- Cached input
- $0.025
- batch $0.05 / $0.2
- Context
- 1.05M
- 33K max output
Where to buy it (3)
Azure AI Foundrydefault region$0.1 / $0.4cached $0.03
OpenAI (list)the lab's own API$0.1 / $0.4cached $0.025
Azure AI Foundryswedencentral+10%$0.11 / $0.44cached $0.033
Through a gateway
Your own numbers100M input and 20M output tokens a month: $18.00 direct.
Helicone$18.00+0%
Kong AI Gateway$18.00+0%
LiteLLM$18.00+0%
Vercel AI Gateway$18.00+0%
AWS Bedrock$18.00+0%
Azure AI Foundry$18.00+0%
Google Vertex AI$18.00+0%
Databricks Mosaic AI$18.00+0%
Hugging Face Inference$18.00+0%
Cloudflare AI Gateway$18.90+5%
Requesty$18.90+5%
Eden AI$18.99+5.5%
OpenRouter$18.99+5.5%
Snowflake Cortex AI$21.60+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
GPT-6 LunaOpenAI · 1.05M$0.1 / $0.5in / out per 1M$0.1$0.5$0.011.05MSep 22, 2026
GPT-6 Luna ProOpenAI · 1.05M$0.1 / $0.5in / out per 1M$0.1$0.5$0.011.05MSep 22, 2026
GPT-6 SolOpenAI · 1.05M$2 / $10in / out per 1M$2$10$0.21.05MSep 22, 2026
GPT-6 Sol ProOpenAI · 1.05M$2 / $10in / out per 1M$2$10$0.21.05MSep 22, 2026
DeepSeek V4.1 FlashDeepSeek · 1M$0.15 / $0.6in / out per 1M$0.15$0.6$0.0031MSep 10, 2026
Qwen3.8 Omni FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MSep 21, 2026
Muse Spark 1.3 ContributorMeta · 1M$0.1 / $0.2in / out per 1M$0.1$0.2$0.0021MSep 2, 2026
Qwen3.8 FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MAug 26, 2026
About GPT-4.1 Nano
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series.
Takes image, text, files in and returns text. First listed Apr 15, 2025. Blended price $0.175 per 1M tokens at 3 input to 1 output.
Questions
How much does GPT-4.1 Nano cost?
$0.1 per million input tokens and $0.4 per million output tokens, with cached input at $0.025, or $0.05 and $0.2 through the Batch API.
What is the context window of GPT-4.1 Nano?
1.05M tokens, with up to 33K tokens of output.
Where is GPT-4.1 Nano cheapest?
Of 3 routes tracked, Azure AI Foundry is cheapest at $0.1 input and $0.4 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.