API pricing per million tokens
- Input
- $1.2
- per 1M, list
- Output
- $4
- per 1M tokens
- Cached input
- $0.24
- no batch tier listed
- Context
- 203K
- 131K max output
Where to buy it (2)
Z.ai (list)the lab's own API$1.2 / $4cached $0.24
Z.aifp8, fp8$1.2 / $4cached $0.24
Through a gateway
Your own numbers100M input and 20M output tokens a month: $200.00 direct.
Helicone$200.00+0%
Kong AI Gateway$200.00+0%
LiteLLM$200.00+0%
Vercel AI Gateway$200.00+0%
AWS Bedrock$200.00+0%
Azure AI Foundry$200.00+0%
Google Vertex AI$200.00+0%
Databricks Mosaic AI$200.00+0%
Hugging Face Inference$200.00+0%
Cloudflare AI Gateway$210.00+5%
Requesty$210.00+5%
Eden AI$211.00+5.5%
OpenRouter$211.00+5.5%
Snowflake Cortex AI$240.00+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
GLM 5.3 PrimeZ.ai · 1M$2.8 / $8.8in / out per 1M$2.8$8.8$0.561MSep 23, 2026
GLM 5.3 FlashXZ.ai · 1M$0.37 / $1.25in / out per 1M$0.37$1.25$0.091MSep 18, 2026
GLM 5.3 FlashZ.ai · via InferenceNet (fp4) · 1.31M$0.045 / $0.14in / out per 1M$0.045$0.14$0.011.31MAug 26, 2026
GLM 5.3Z.ai · via DeepInfra (fp4) · 1.31M$0.563 / $2.5in / out per 1M$0.563$2.5$0.1251.31MAug 18, 2026
GPT-5.6 SolOpenAI · 1.05M$2 / $10in / out per 1M$2$10$0.21.05MJul 9, 2026
Muse Spark 1.3Meta · 1M$1.25 / $4.25in / out per 1M$1.25$4.25$0.151MSep 2, 2026
Gemini 3.8 FlashGoogle Gemini API · 1M$0.75 / $3.75in / out per 1M$0.75$3.75$0.0751MSep 2, 2026
Muse Spark 1.1Meta · 1M$1.25 / $4.25in / out per 1M$1.25$4.25$0.151MJul 16, 2026
About GLM 5V Turbo
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.
Takes image, text, video in and returns text. First listed Apr 2, 2026. Blended price $1.9 per 1M tokens at 3 input to 1 output.
Questions
How much does GLM 5V Turbo cost?
$1.2 per million input tokens and $4 per million output tokens, with cached input at $0.24.
What is the context window of GLM 5V Turbo?
203K tokens, with up to 131K tokens of output.
Where is GLM 5V Turbo cheapest?
Of 2 routes tracked, Z.ai is cheapest at $1.2 input and $4 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.