API pricing per million tokens
- Input
- $0.163
- per 1M, list
- Output
- $1.3
- per 1M tokens
- Cached input
- n/a
- no batch tier listed
- Context
- 262K
- 16K max output
Where to buy it (7)
- Darkbloomfp4, fp4-45%$0.08 / $0.75$0.248 blended
DeepInfrafp8, fp8-21%$0.14 / $1cached $0.05
Parasailfp8, fp8-19%$0.15 / $1cached $0.05
Qwen (Alibaba) (list)the lab's own API$0.163 / $1.3$0.447 blended
Venicedefault region+22%$0.313 / $1.25cached $0.156
- AtlasCloudfp8, fp8+38%$0.225 / $1.8cached $0.225
SiliconFlowfp8, fp8+41%$0.24 / $1.8cached $0.15
Through a gateway
Your own numbers100M input and 20M output tokens a month: $42.25 direct.
Helicone$42.25+0%
Kong AI Gateway$42.25+0%
LiteLLM$42.25+0%
Vercel AI Gateway$42.25+0%
AWS Bedrock$42.25+0%
Azure AI Foundry$42.25+0%
Google Vertex AI$42.25+0%
Databricks Mosaic AI$42.25+0%
Hugging Face Inference$42.25+0%
Cloudflare AI Gateway$44.36+5%
Requesty$44.36+5%
Eden AI$44.57+5.5%
OpenRouter$44.57+5.5%
Snowflake Cortex AI$50.70+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
Qwen3.8 Max PrimeQwen (Alibaba) · 1M$4 / $12in / out per 1M$4$12$0.51MSep 23, 2026
Qwen3.8 Omni FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MSep 21, 2026
Qwen3.8 Max (0902)Qwen (Alibaba) · 1M$2 / $6in / out per 1M$2$6$0.251MSep 3, 2026
Qwen3.8 FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MAug 26, 2026
GLM 4.6Z.ai · via Venice (fp4) · 205K$0.43 / $1.75in / out per 1M$0.43$1.75$0.08205KSep 30, 2025
DeepSeek V4.1 FlashDeepSeek · 1M$0.15 / $0.6in / out per 1M$0.15$0.6$0.0031MSep 10, 2026
Kimi K2.5Moonshot AI · via SiliconFlow (int4) · 262K$0.45 / $2.25in / out per 1M$0.45$2.25$0.07262KJan 27, 2026
Gemini 3.1 Flash LiteGoogle Gemini API · 1M$0.25 / $1.5in / out per 1M$0.25$1.5$0.0251MMay 7, 2026
About Qwen3.5-35B-A3B
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency.
Takes text, image, video in and returns text. First listed Sep 23, 2026. Blended price $0.447 per 1M tokens at 3 input to 1 output.
Questions
How much does Qwen3.5-35B-A3B cost?
$0.163 per million input tokens and $1.3 per million output tokens.
What is the context window of Qwen3.5-35B-A3B?
262K tokens, with up to 16K tokens of output.
Where is Qwen3.5-35B-A3B cheapest?
Of 7 routes tracked, Darkbloom is cheapest at $0.08 input and $0.75 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.