Skip to content

Compare model prices

Input, output, cached and batch prices, every route and a month of tokens, side by side.

GLM 4.6GLM 5.3 Prime

A month of tokens

  • GLM 4.6$78.00
  • GLM 5.3 Prime$456.00

100M input and 20M output tokens, list price. Use your own numbers.

Side by side

GLM 4.6GLM 5.3 Prime
Input per 1M$0.43$2.8
Output per 1M$1.75$8.8
Cached input$0.08$0.56
Batch in / outn/an/a
Blended (3:1)$0.76$4.3
Context205K1M
Routes42
ReleasedSep 30, 2025Sep 23, 2026
Leadersboard score5 (#30)n/a

On the price map

Size: context window, 8K to 2MEqual blended cost (3:1) 243 models, lower left is cheaper

Every route

GLM 4.6

  • Venicefp4, fp4$0.43 / $1.75cached $0.08
  • DeepInfrafp4, fp4+15%$0.5 / $2cached $0.1
  • Novita AIbf16, bf16+27%$0.55 / $2.2cached $0.11
  • Z.aifp4, fp4+32%$0.6 / $2.2cached $0.11

GLM 5.3 Prime

  • Z.ai (list)the lab's own API$2.8 / $8.8cached $0.56
  • Qwen (Alibaba)default region$2.8 / $8.8cached $0.56

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.