API pricing per million tokens
- Input
- $0.26
- per 1M, list
- Output
- $2.08
- per 1M tokens
- Cached input
- n/a
- no batch tier listed
- Context
- 262K
- 66K max output
Where to buy it (5)
Qwen (Alibaba) (list)the lab's own API$0.26 / $2.08$0.715 blended
SiliconFlowfp8, fp8$0.26 / $2.08$0.715 blended
DeepInfrafp4, fp4+14%$0.29 / $2.4$0.817 blended
- AtlasCloudfp8, fp8+15%$0.3 / $2.4cached $0.3
Novita AIbf16, bf16+54%$0.4 / $3.2$1.1 blended
Through a gateway
Your own numbers100M input and 20M output tokens a month: $67.60 direct.
Helicone$67.60+0%
Kong AI Gateway$67.60+0%
LiteLLM$67.60+0%
Vercel AI Gateway$67.60+0%
AWS Bedrock$67.60+0%
Azure AI Foundry$67.60+0%
Google Vertex AI$67.60+0%
Databricks Mosaic AI$67.60+0%
Hugging Face Inference$67.60+0%
Cloudflare AI Gateway$70.98+5%
Requesty$70.98+5%
Eden AI$71.32+5.5%
OpenRouter$71.32+5.5%
Snowflake Cortex AI$81.12+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
Qwen3.8 Max PrimeQwen (Alibaba) · 1M$4 / $12in / out per 1M$4$12$0.51MSep 23, 2026
Qwen3.8 Omni FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MSep 21, 2026
Qwen3.8 Max (0902)Qwen (Alibaba) · 1M$2 / $6in / out per 1M$2$6$0.251MSep 3, 2026
Qwen3.8 FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MAug 26, 2026
Gemini 3.8 FlashGoogle Gemini API · 1M$0.75 / $3.75in / out per 1M$0.75$3.75$0.0751MSep 2, 2026
GLM 5.3Z.ai · via DeepInfra (fp4) · 1.31M$0.563 / $2.5in / out per 1M$0.563$2.5$0.1251.31MAug 18, 2026
Gemini 3.7 FlashGoogle Gemini API · 1M$0.75 / $3.75in / out per 1M$0.75$3.75$0.0751MAug 13, 2026
GLM 4.6Z.ai · via Venice (fp4) · 205K$0.43 / $1.75in / out per 1M$0.43$1.75$0.08205KSep 30, 2025
About Qwen3.5-122B-A10B
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency.
Takes text, image, video in and returns text. First listed Sep 23, 2026. Blended price $0.715 per 1M tokens at 3 input to 1 output.
Questions
How much does Qwen3.5-122B-A10B cost?
$0.26 per million input tokens and $2.08 per million output tokens.
What is the context window of Qwen3.5-122B-A10B?
262K tokens, with up to 66K tokens of output.
Where is Qwen3.5-122B-A10B cheapest?
Of 5 routes tracked, Qwen (Alibaba) is cheapest at $0.26 input and $2.08 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.