Skip to content

Qwen3 235B A22B Instruct 2507

API pricing per million tokens

qwen3-235b-a22b-2507 Open weightsChecked 6h ago
Cost calculator Compare
Input
$0.15
per 1M, list
Output
$0.598
per 1M tokens
Cached input
n/a
no batch tier listed
Context
262K
236K max output

Where to buy it (9)

  • GMICloudfp8, fp8-41%$0.087 / $0.35cached $0.018
  • DeepInfrafp8, fp8-22%$0.09 / $0.55$0.205 blended
  • Novita AIfp8, fp8-19%$0.09 / $0.58$0.213 blended
  • Qwen (Alibaba) (list)the lab's own API$0.15 / $0.598$0.262 blended
  • Venicefp8, fp8+15%$0.15 / $0.75$0.3 blended
  • Nebiusfp8, fp8+15%$0.2 / $0.6$0.3 blended
  • Parasailfp8, fp8+17%$0.14 / $0.8cached $0.05
  • StreamLakedefault region+40%$0.21 / $0.84$0.368 blended
  • Google Vertex AIus-south1+47%$0.22 / $0.88$0.385 blended

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $26.91 direct.

Related models

About Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass.

Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.262 per 1M tokens at 3 input to 1 output.

Across fru.devCompany profile

Questions

How much does Qwen3 235B A22B Instruct 2507 cost?

$0.15 per million input tokens and $0.598 per million output tokens.

What is the context window of Qwen3 235B A22B Instruct 2507?

262K tokens, with up to 236K tokens of output.

Where is Qwen3 235B A22B Instruct 2507 cheapest?

Of 9 routes tracked, GMICloud is cheapest at $0.087 input and $0.35 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.