Ways to buy 2
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
The lab's own API · cached $0.09 | $0.37 | $1.25 | $0.37 / $1.25List price | $0.09 | List price |
fp8 · cached $0.09 | $0.37 | $1.25 | $0.37 / $1.25Same as list | $0.09 | Same as list |
USD per 1M tokens. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $62.00 direct.
Helicone$62.00+0%
Kong AI Gateway$62.00+0%
LiteLLM$62.00+0%
Vercel AI Gateway$62.00+0%
Cloudflare AI Gateway$65.10+5%
Requesty$65.10+5%
Eden AI$65.41+5.5%
OpenRouter$65.41+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
Z.ai · 1M | $2.8 | $8.8 | $2.8 / $8.8 | $4.3 | 1M | Lab list price | · |
Z.ai · 1.31M | $0.045 | $0.14 | $0.045 / $0.14 | $0.069 | 1.31M | on InferenceNet (fp4) | · |
Z.ai · 1.31M | $0.563 | $2.5 | $0.563 / $2.5 | $1.05 | 1.31M | on DeepInfra (fp4) | · |
Z.ai · 1M | $0.563 | $1.8 | $0.563 / $1.8 | $0.872 | 1M | on DeepInfra (fp4) | · |
Moonshot AI · 262K | $0.45 | $2.25 | $0.45 / $2.25 | $0.9 | 262K | on SiliconFlow (int4) | · |
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.281 on Google Gemini API (flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Gemini API (flex) | · |
Cohere · 192K | $0.3 | $1.5 | $0.3 / $1.5 | $0.6 | 192K | Lab list price | · |
About GLM 5.3 FlashX
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
Takes text, image, video in and returns text. First listed Sep 19, 2026. Up to 131K output tokens. Blended price $0.59 per 1M tokens at 3 input to 1 output. Model id glm-5-3-flashx.
Price history
List price since Sep 19, 2026, from Internet Archive snapshots of the OpenRouter catalog, then daily reads.
Across fru.devCompany profile
Questions
How much does GLM 5.3 FlashX cost?
$0.37 per million input tokens and $1.25 per million output tokens, with cached input at $0.09.
What is the context window of GLM 5.3 FlashX?
1M tokens, with up to 131K tokens of output.
Where is GLM 5.3 FlashX cheapest?
Of 2 routes tracked, Z.ai is cheapest at $0.37 input and $1.25 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.