API pricing per million tokens
- Input
- $0.75
- per 1M, list
- Output
- $3.75
- per 1M tokens
- Cached input
- $0.075
- batch $0.375 / $1.88
- Context
- 1M
- 66K max output
Where to buy it (6)
Google Vertex AIglobal-flex-50%$0.375 / $1.88cached $0.037
Google Gemini APIflex-50%$0.375 / $1.88cached $0.037
Google Vertex AIglobal$0.75 / $3.75cached $0.075
Google Gemini API (list)the lab's own API$0.75 / $3.75cached $0.075
Google Vertex AIglobal-priority+80%$1.35 / $6.75cached $0.135
Google Gemini APIpriority+80%$1.35 / $6.75cached $0.135
Through a gateway
Your own numbers100M input and 20M output tokens a month: $150.00 direct.
Helicone$150.00+0%
Kong AI Gateway$150.00+0%
LiteLLM$150.00+0%
Vercel AI Gateway$150.00+0%
AWS Bedrock$150.00+0%
Azure AI Foundry$150.00+0%
Google Vertex AI$150.00+0%
Databricks Mosaic AI$150.00+0%
Hugging Face Inference$150.00+0%
Cloudflare AI Gateway$157.50+5%
Requesty$157.50+5%
Eden AI$158.25+5.5%
OpenRouter$158.25+5.5%
Snowflake Cortex AI$180.00+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
Gemini 3.7 FlashGoogle Gemini API · 1M$0.75 / $3.75in / out per 1M$0.75$3.75$0.0751MAug 13, 2026
Gemini 3.5 Flash LiteGoogle Gemini API · 1M$0.3 / $2.5in / out per 1M$0.3$2.5$0.031MJul 21, 2026
Gemini 3.6 FlashGoogle Gemini API · 1M$0.75 / $3.75in / out per 1M$0.75$3.75$0.0751MJul 21, 2026
Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google Gemini API · 66K$0.25 / $1.5in / out per 1M$0.25$1.5n/a66KJun 30, 2026
Muse Spark 1.3Meta · 1M$1.25 / $4.25in / out per 1M$1.25$4.25$0.151MSep 2, 2026
Muse Spark 1.1Meta · 1M$1.25 / $4.25in / out per 1M$1.25$4.25$0.151MJul 16, 2026
Qwen3.8 Max (0902)Qwen (Alibaba) · 1M$2 / $6in / out per 1M$2$6$0.251MSep 3, 2026
GLM 5.3Z.ai · via DeepInfra (fp4) · 1.31M$0.563 / $2.5in / out per 1M$0.563$2.5$0.1251.31MAug 18, 2026
About Gemini 3.8 Flash
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Takes text, image, video, files, audio in and returns text. First listed Sep 2, 2026. Blended price $1.5 per 1M tokens at 3 input to 1 output. Overall score 15 on leadersboard.fru.dev (rank 13).
Questions
How much does Gemini 3.8 Flash cost?
$0.75 per million input tokens and $3.75 per million output tokens, with cached input at $0.075, or $0.375 and $1.88 through the Batch API.
What is the context window of Gemini 3.8 Flash?
1M tokens, with up to 66K tokens of output.
Where is Gemini 3.8 Flash cheapest?
Of 6 routes tracked, Google Vertex AI is cheapest at $0.375 input and $1.88 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.