Skip to content

GLM 4.6

API pricing per million tokens

glm-4-6 Open weightsChecked 5h ago
Input
$0.43
per 1M, Venice (fp4)
Output
$1.75
per 1M tokens
Cached input
$0.08
no batch tier listed
Context
205K
16K max output

Where to buy it (4)

  • Venicefp4, fp4$0.43 / $1.75cached $0.08
  • DeepInfrafp4, fp4$0.5 / $2cached $0.1
  • Novita AIbf16, bf16$0.55 / $2.2cached $0.11
  • Z.aifp4, fp4$0.6 / $2.2cached $0.11

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $78.00 direct.

Related models

About GLM 4.6

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.76 per 1M tokens at 3 input to 1 output. Overall score 5 on leadersboard.fru.dev (rank 30).

Across fru.devCompany profile

Questions

How much does GLM 4.6 cost?

$0.43 per million input tokens and $1.75 per million output tokens, with cached input at $0.08.

What is the context window of GLM 4.6?

205K tokens, with up to 16K tokens of output.

Where is GLM 4.6 cheapest?

Of 4 routes tracked, Venice is cheapest at $0.43 input and $1.75 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.