Compare model prices
Input, output, cached and batch prices, every route and a month of tokens, side by side.
A month of tokens
GLM 4.6$78.00
GLM 5.3 Prime$456.00
100M input and 20M output tokens, list price. Use your own numbers.
Side by side
| GLM 4.6 | GLM 5.3 Prime | |
|---|---|---|
| Input per 1M | $0.43 | $2.8 |
| Output per 1M | $1.75 | $8.8 |
| Cached input | $0.08 | $0.56 |
| Batch in / out | n/a | n/a |
| Blended (3:1) | $0.76 | $4.3 |
| Context | 205K | 1M |
| Routes | 4 | 2 |
| Released | Sep 30, 2025 | Sep 23, 2026 |
| Leadersboard score | 5 (#30) | n/a |
On the price map
Size: context window, 8K to 2MEqual blended cost (3:1) 243 models, lower left is cheaper
Every route
GLM 4.6
Venicefp4, fp4$0.43 / $1.75cached $0.08
DeepInfrafp4, fp4+15%$0.5 / $2cached $0.1
Novita AIbf16, bf16+27%$0.55 / $2.2cached $0.11
Z.aifp4, fp4+32%$0.6 / $2.2cached $0.11
GLM 5.3 Prime
Z.ai (list)the lab's own API$2.8 / $8.8cached $0.56
Qwen (Alibaba)default region$2.8 / $8.8cached $0.56
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.