API pricing per million tokens
- Input
- $0.117
- per 1M, list
- Output
- $0.455
- per 1M tokens
- Cached input
- n/a
- no batch tier listed
- Context
- 262K
- 33K max output
Where to buy it (2)
Qwen (Alibaba) (list)the lab's own API$0.117 / $0.455$0.202 blended
Parasailbf16, bf16+86%$0.25 / $0.75cached $0.12
Through a gateway
Your own numbers100M input and 20M output tokens a month: $20.80 direct.
Helicone$20.80+0%
Kong AI Gateway$20.80+0%
LiteLLM$20.80+0%
Vercel AI Gateway$20.80+0%
AWS Bedrock$20.80+0%
Azure AI Foundry$20.80+0%
Google Vertex AI$20.80+0%
Databricks Mosaic AI$20.80+0%
Hugging Face Inference$20.80+0%
Cloudflare AI Gateway$21.84+5%
Requesty$21.84+5%
Eden AI$21.94+5.5%
OpenRouter$21.94+5.5%
Snowflake Cortex AI$24.96+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
Qwen3.8 Max PrimeQwen (Alibaba) · 1M$4 / $12in / out per 1M$4$12$0.51MSep 23, 2026
Qwen3.8 Omni FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MSep 21, 2026
Qwen3.8 Max (0902)Qwen (Alibaba) · 1M$2 / $6in / out per 1M$2$6$0.251MSep 3, 2026
Qwen3.8 FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MAug 26, 2026
DeepSeek V4.1 FlashDeepSeek · 1M$0.15 / $0.6in / out per 1M$0.15$0.6$0.0031MSep 10, 2026
GPT-6 LunaOpenAI · 1.05M$0.1 / $0.5in / out per 1M$0.1$0.5$0.011.05MSep 22, 2026
GPT-6 Luna ProOpenAI · 1.05M$0.1 / $0.5in / out per 1M$0.1$0.5$0.011.05MSep 22, 2026
Muse Spark 1.3 ContributorMeta · 1M$0.1 / $0.2in / out per 1M$0.1$0.2$0.0021MSep 2, 2026
About Qwen3 VL 8B Instruct
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video.
Takes image, text in and returns text. First listed Sep 23, 2026. Blended price $0.202 per 1M tokens at 3 input to 1 output.
Questions
How much does Qwen3 VL 8B Instruct cost?
$0.117 per million input tokens and $0.455 per million output tokens.
What is the context window of Qwen3 VL 8B Instruct?
262K tokens, with up to 33K tokens of output.
Where is Qwen3 VL 8B Instruct cheapest?
Of 2 routes tracked, Qwen (Alibaba) is cheapest at $0.117 input and $0.455 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.