Ways to buy 5
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
flex, cheapest · cached $0.005 | $0.05 | $0.2 | $0.05 / $0.2-50% vs list | $0.005 | -50% vs list |
The lab's own API · cached $0.01 | $0.1 | $0.4 | $0.1 / $0.4List price | $0.01 | List price |
Default region · cached $0.01 | $0.1 | $0.4 | $0.1 / $0.4Same as list | $0.01 | Same as list |
eu · cached $0.01 | $0.1 | $0.4 | $0.1 / $0.4Same as list | $0.01 | Same as list |
priority · cached $0.018 | $0.18 | $0.72 | $0.18 / $0.72+80% vs list | $0.018 | +80% vs list |
USD per 1M tokens; batch $0.05 in, $0.2 out on the lab's API. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $18.00 direct.
Helicone$18.00+0%
Kong AI Gateway$18.00+0%
LiteLLM$18.00+0%
Vercel AI Gateway$18.00+0%
Cloudflare AI Gateway$18.90+5%
Requesty$18.90+5%
Eden AI$18.99+5.5%
OpenRouter$18.99+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Google · 1MCheapest: $0.75 on Google Gemini API (flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Gemini API (flex) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75100% | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | 100% |
Google · 1MCheapest: $0.425 on Google Vertex AI (global-flex) | $0.3 | $2.5 | $0.3 / $2.5 | $0.85 | 1M | $0.425 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.75 on Google Vertex AI (global-flex) | $0.75 | $3.75 | $0.75 / $3.75 | $1.5 | 1M | $0.75 on Google Vertex AI (global-flex) | · |
DeepSeek · 1MCheapest: $0.1 on Relace | $0.15 | $0.6 | $0.15 / $0.6 | $0.262 | 1M | $0.1 on Relace | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
Qwen (Alibaba) · 1M | $0.15 | $0.47 | $0.15 / $0.47 | $0.23 | 1M | Lab list price | · |
About Gemini 2.5 Flash Lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.
Takes text, image, files, audio, video in and returns text. First listed Jul 24, 2025. Up to 66K output tokens. Blended price $0.175 per 1M tokens at 3 input to 1 output. Model id gemini-2-5-flash-lite.
Price history
List price since Jul 24, 2025, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.
| Seen | Model | Old price | New price | Change |
|---|---|---|---|---|
| Oct 20, 2025 | Cached input $0.025 | Cached input $0.01 | -60% |
Across fru.devCompany profile7 acquisitionsPay schedule7 conferences
Questions
How much does Gemini 2.5 Flash Lite cost?
$0.1 per million input tokens and $0.4 per million output tokens, with cached input at $0.01, or $0.05 and $0.2 through the Batch API.
What is the context window of Gemini 2.5 Flash Lite?
1M tokens, with up to 66K tokens of output.
Where is Gemini 2.5 Flash Lite cheapest?
Of 5 routes tracked, Google Gemini API is cheapest at $0.05 input and $0.2 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.