Skip to content

GLM 4.7

API pricing per million tokens

glm-4-7 Open weightsChecked 5h ago
Cost calculator Compare
Input
$0.4
per 1M, DeepInfra (fp4)
Output
$1.75
per 1M tokens
Cached input
$0.08
no batch tier listed
Context
205K
131K max output

Where to buy it (7)

  • DeepInfrafp4, fp4$0.4 / $1.75cached $0.08
  • Venicefp4, fp4$0.4 / $1.93cached $0.08
  • AtlasCloudfp8, fp8$0.52 / $1.85cached $0.12
  • Novita AIfp8, fp8$0.54 / $1.98cached $0.099
  • Google Vertex AIdefault region$0.6 / $2.2$1 blended
  • Z.aifp4, fp4$0.6 / $2.2cached $0.11
  • Mancer 2fp4, fp4$0.7 / $2.5$1.15 blended

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $75.00 direct.

Related models

About GLM 4.7

GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution.

Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.738 per 1M tokens at 3 input to 1 output.

Across fru.devCompany profile

Questions

How much does GLM 4.7 cost?

$0.4 per million input tokens and $1.75 per million output tokens, with cached input at $0.08.

What is the context window of GLM 4.7?

205K tokens, with up to 131K tokens of output.

Where is GLM 4.7 cheapest?

Of 7 routes tracked, DeepInfra is cheapest at $0.4 input and $1.75 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.