Skip to content

API pricing per million tokens

qwen3-32b Open weightsChecked 7h ago
Cost calculator Compare
Input
$0.08
per 1M, DeepInfra (fp8)
Output
$0.28
per 1M tokens
Cached input
n/a
no batch tier listed
Context
131K
16K max output

Where to buy it (2)

  • DeepInfrafp8, fp8$0.08 / $0.28$0.13 blended
  • SiliconFlowfp8, fp8$0.14 / $0.57$0.248 blended

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $13.60 direct.

Related models

About Qwen3 32B

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue.

Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.13 per 1M tokens at 3 input to 1 output.

Across fru.devCompany profile

Questions

How much does Qwen3 32B cost?

$0.08 per million input tokens and $0.28 per million output tokens.

What is the context window of Qwen3 32B?

131K tokens, with up to 16K tokens of output.

Where is Qwen3 32B cheapest?

Of 2 routes tracked, DeepInfra is cheapest at $0.08 input and $0.28 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.