API pricing per million tokens
- Input
- $0.425
- per 1M, list
- Output
- $2.55
- per 1M tokens
- Cached input
- $0.085
- no batch tier listed
- Context
- 1M
- 131K max output
Where to buy it (16)
- Darkbloomfp4, fp4-45%$0.1 / $1.8$0.525 blended
DeepInfrabf16, bf16-39%$0.15 / $1.88cached $0.037
- Phaladefault region-30%$0.199 / $2.08cached $0.042
- DekaLLMdefault region-27%$0.1 / $2.5cached $0.05
Chutesfp8, fp8-24%$0.24 / $2.2cached $0.024
Parasailfp8, fp8-24%$0.24 / $2.2cached $0.05
- AkashMLfp8, fp8-23%$0.25 / $2.2cached $0.05
- Mancer 2fp8, fp8-19%$0.2 / $2.5$0.775 blended
- Ionstreamfp8, fp8-11%$0.28 / $2.55cached $0.1
Qwen (Alibaba) (list)the lab's own API$0.425 / $2.55cached $0.085
CoreWeavefp8, fp8+10%$0.4 / $3cached $0.15
Novita AIdefault region+11%$0.42 / $3cached $0.085
Cloudflare Workers AIdefault region+19%$0.45 / $3.2cached $0.05
Venicefp8, fp8+19%$0.45 / $3.2$1.14 blended
- Waferdefault region+23%$0.097 / $4.4cached $0.087
- Rekafp8, fp8+23%$0.098 / $4.4cached $0.089
Through a gateway
Your own numbers100M input and 20M output tokens a month: $93.50 direct.
Helicone$93.50+0%
Kong AI Gateway$93.50+0%
LiteLLM$93.50+0%
Vercel AI Gateway$93.50+0%
AWS Bedrock$93.50+0%
Azure AI Foundry$93.50+0%
Google Vertex AI$93.50+0%
Databricks Mosaic AI$93.50+0%
Hugging Face Inference$93.50+0%
Cloudflare AI Gateway$98.17+5%
Requesty$98.17+5%
Eden AI$98.64+5.5%
OpenRouter$98.64+5.5%
Snowflake Cortex AI$112.20+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
Qwen3.8 Max PrimeQwen (Alibaba) · 1M$4 / $12in / out per 1M$4$12$0.51MSep 23, 2026
Qwen3.8 Omni FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MSep 21, 2026
Qwen3.8 Max (0902)Qwen (Alibaba) · 1M$2 / $6in / out per 1M$2$6$0.251MSep 3, 2026
Qwen3.8 FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MAug 26, 2026
Muse Spark 1.3Meta · 1M$1.25 / $4.25in / out per 1M$1.25$4.25$0.151MSep 2, 2026
Gemini 3.8 FlashGoogle Gemini API · 1M$0.75 / $3.75in / out per 1M$0.75$3.75$0.0751MSep 2, 2026
Muse Spark 1.1Meta · 1M$1.25 / $4.25in / out per 1M$1.25$4.25$0.151MJul 16, 2026
GLM 5.3Z.ai · via DeepInfra (fp4) · 1.31M$0.563 / $2.5in / out per 1M$0.563$2.5$0.1251.31MAug 18, 2026
About Qwen3.8 27B
Qwen3.8 27B is an open-weight dense vision-language model from Qwen.
Takes text, image, video in and returns text. First listed Sep 23, 2026. Blended price $0.956 per 1M tokens at 3 input to 1 output.
Questions
How much does Qwen3.8 27B cost?
$0.425 per million input tokens and $2.55 per million output tokens, with cached input at $0.085.
What is the context window of Qwen3.8 27B?
1M tokens, with up to 131K tokens of output.
Where is Qwen3.8 27B cheapest?
Of 16 routes tracked, Darkbloom is cheapest at $0.1 input and $1.8 output per million tokens.
Has Qwen3.8 27B's price changed?
Yes. The input price went from $0.121 to $0.099, first seen Sep 24, 2026.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.