API pricing per million tokens
- Input
- $0.13
- per 1M, Novita AI (bf16)
- Output
- $0.85
- per 1M tokens
- Cached input
- $0.025
- no batch tier listed
- Context
- 131K
- 98K max output
Where to buy it (3)
Novita AIbf16, bf16$0.13 / $0.85cached $0.025
SiliconFlowfp8, fp8$0.14 / $0.86$0.32 blended
Z.aifp8, fp8$0.2 / $1.1cached $0.03
Through a gateway
Your own numbers100M input and 20M output tokens a month: $30.00 direct.
Helicone$30.00+0%
Kong AI Gateway$30.00+0%
LiteLLM$30.00+0%
Vercel AI Gateway$30.00+0%
AWS Bedrock$30.00+0%
Azure AI Foundry$30.00+0%
Google Vertex AI$30.00+0%
Databricks Mosaic AI$30.00+0%
Hugging Face Inference$30.00+0%
Cloudflare AI Gateway$31.50+5%
Requesty$31.50+5%
Eden AI$31.65+5.5%
OpenRouter$31.65+5.5%
Snowflake Cortex AI$36.00+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
GLM 5.3 PrimeZ.ai · 1M$2.8 / $8.8in / out per 1M$2.8$8.8$0.561MSep 23, 2026
GLM 5.3 FlashXZ.ai · 1M$0.37 / $1.25in / out per 1M$0.37$1.25$0.091MSep 18, 2026
GLM 5.3 FlashZ.ai · via InferenceNet (fp4) · 1.31M$0.045 / $0.14in / out per 1M$0.045$0.14$0.011.31MAug 26, 2026
GLM 5.3Z.ai · via DeepInfra (fp4) · 1.31M$0.563 / $2.5in / out per 1M$0.563$2.5$0.1251.31MAug 18, 2026
DeepSeek V4.1 FlashDeepSeek · 1M$0.15 / $0.6in / out per 1M$0.15$0.6$0.0031MSep 10, 2026
Gemini 3.1 Flash LiteGoogle Gemini API · 1M$0.25 / $1.5in / out per 1M$0.25$1.5$0.0251MMay 7, 2026
Gemini 3.1 Flash Lite PreviewGoogle Gemini API · 1M$0.25 / $1.5in / out per 1M$0.25$1.5$0.0251MMar 3, 2026
Command A+Cohere · 192K$0.3 / $1.5in / out per 1M$0.3$1.5$0.15192KSep 22, 2026
About GLM 4.5 Air
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications.
Takes text in and returns text. First listed Sep 23, 2026. Blended price $0.31 per 1M tokens at 3 input to 1 output.
Questions
How much does GLM 4.5 Air cost?
$0.13 per million input tokens and $0.85 per million output tokens, with cached input at $0.025.
What is the context window of GLM 4.5 Air?
131K tokens, with up to 98K tokens of output.
Where is GLM 4.5 Air cheapest?
Of 3 routes tracked, Novita AI is cheapest at $0.13 input and $0.85 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.