Skip to content

Qwen2.5 72B Instruct

API pricing per million tokens

qwen-2-5-72b-instruct Open weightsChecked 8h ago
Cost calculator Compare
Input
$0.36
per 1M, DeepInfra (fp8)
Output
$0.4
per 1M tokens
Cached input
n/a
no batch tier listed
Context
33K
16K max output

Where to buy it (2)

  • DeepInfrafp8, fp8$0.36 / $0.4$0.37 blended
  • Novita AIbf16, bf16$0.38 / $0.4$0.385 blended

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $44.00 direct.

Related models

About Qwen2.5 72B Instruct

Qwen2.5 72B is the latest series of Qwen large language models.

Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.37 per 1M tokens at 3 input to 1 output.

Across fru.devCompany profile

Questions

How much does Qwen2.5 72B Instruct cost?

$0.36 per million input tokens and $0.4 per million output tokens.

What is the context window of Qwen2.5 72B Instruct?

33K tokens, with up to 16K tokens of output.

Where is Qwen2.5 72B Instruct cheapest?

Of 2 routes tracked, DeepInfra is cheapest at $0.36 input and $0.4 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.