Skip to content

GLM 4.7 Flash

API pricing per million tokens

glm-4-7-flash Open weightsChecked 7h ago
Cost calculator Compare
Input
$0.06
per 1M, Venice (fp8)
Output
$0.4
per 1M tokens
Cached input
$0.01
no batch tier listed
Context
200K
118K max output

Where to buy it (3)

  • Venicefp8, fp8$0.06 / $0.4cached $0.01
  • Cloudflare Workers AIdefault region$0.06 / $0.4$0.145 blended
  • Novita AIbf16, bf16$0.07 / $0.4cached $0.01

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $14.00 direct.

Related models

About GLM 4.7 Flash

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.

Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.145 per 1M tokens at 3 input to 1 output.

Across fru.devCompany profile

Questions

How much does GLM 4.7 Flash cost?

$0.06 per million input tokens and $0.4 per million output tokens, with cached input at $0.01.

What is the context window of GLM 4.7 Flash?

200K tokens, with up to 118K tokens of output.

Where is GLM 4.7 Flash cheapest?

Of 3 routes tracked, Venice is cheapest at $0.06 input and $0.4 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.