Back to models
available

Z.ai: GLM 5.3

z-ai/glm-5-3

Prices

Input

$1.40 per 1M tokens

low
Input (cache write)

$0.00 per 1M tokens

low
Input (cached read)

$0.26 per 1M tokens

low
Output

$4.40 per 1M tokens

low

Also available on

The same model, sold by someone else. The price above is the maker's own.

37 sales channels for this model.
ChannelWhatPrice
DatabricksInput$1.40 per 1M tokens
Input (cache write)$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Cloudflare Workers AIInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Fireworks AI · fireworks-aiInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Alibaba Model StudioInput$1.40 per 1M tokens
Input (cached read)$0.28 per 1M tokens
Output$4.40 per 1M tokens
Alibaba Model Studio · alibaba-cnInput$1.10 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.275 per 1M tokens
Output$3.85 per 1M tokens
OpenRouterInput$1.40 per 1M tokens
Input (cached read)$0.14 per 1M tokens
Output$4.40 per 1M tokens
OpenRouter · z-aiInput$1.40 per 1M tokens
Input (cached read)$0.14 per 1M tokens
Output$4.40 per 1M tokens
Coding plan · volcengine-coding-planInput$0.00 per 1M tokens
Input (cached read)$0.00 per 1M tokens
Output$0.00 per 1M tokens
Coding plan · zai-coding-planInput$0.00 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.00 per 1M tokens
Output$0.00 per 1M tokens
Coding plan · zhipuai-coding-planInput$0.00 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.00 per 1M tokens
Output$0.00 per 1M tokens
Token plan · alibaba-token-planInput$0.00 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.00 per 1M tokens
Output$0.00 per 1M tokens
Token plan · alibaba-token-plan-cnInput$0.00 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.00 per 1M tokens
Output$0.00 per 1M tokens
Other · aiandInput$1.00 per 1M tokens
Input (cached read)$0.30 per 1M tokens
Output$4.00 per 1M tokens
Other · basetenInput$1.40 per 1M tokens
Input (cached read)$0.14 per 1M tokens
Output$4.40 per 1M tokens
Other · crossmodelInput$1.20 per 1M tokens
Input (cache write)$1.20 per 1M tokens
Input (cached read)$0.30 per 1M tokens
Output$4.40 per 1M tokens
Other · deepinfraInput$0.90 per 1M tokens
Input (cached read)$0.20 per 1M tokens
Output$4.00 per 1M tokens
Other · edenaiInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · friendliInput$1.26 per 1M tokens
Input (cached read)$0.234 per 1M tokens
Output$3.96 per 1M tokens
Other · huggingfaceInput$1.40 per 1M tokens
Output$4.40 per 1M tokens
Other · iteracomputeInput$1.20 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$3.50 per 1M tokens
Other · kiloInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · llmgateway-providersInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · merge-gatewayInput$0.70 per 1M tokens
Input (cached read)$0.13 per 1M tokens
Output$2.20 per 1M tokens
Other · nano-gptInput$1.00 per 1M tokens
Input (cached read)$0.20 per 1M tokens
Output$3.20 per 1M tokens
Other · nebiusInput$1.40 per 1M tokens
Input (cached read)$1.40 per 1M tokens
Output$4.40 per 1M tokens
Other · nvidiaInput$0.00 per 1M tokens
Output$0.00 per 1M tokens
Other · ofoxInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · openrouterInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · orcarouterInput$1.26 per 1M tokens
Input (cached read)$0.234 per 1M tokens
Output$3.96 per 1M tokens
Other · siliconflowInput$1.40 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · syntheticInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · temprInput$1.40 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · tensorxInput$1.75 per 1M tokens
Input (cached read)$0.44 per 1M tokens
Output$4.50 per 1M tokens
Other · togetheraiInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · tokengoInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens
Other · vercelInput$1.40 per 1M tokens
Input (cached read)$0.14 per 1M tokens
Output$4.40 per 1M tokens
Other · zenmuxInput$1.40 per 1M tokens
Input (cached read)$0.26 per 1M tokens
Output$4.40 per 1M tokens

Specifications

Provider
z-ai
Context window
1,048,576
Max output tokens
943,718
Open weights
yes
Released
08/14/2026
Retires
no date announced
First seen by this tracker

Price history

No price of this model has changed since we started tracking it.

Input (cache write)
0 per 1M tokens, unchanged since 09/14/2026
Input (cached read)
0.26 per 1M tokens, unchanged since 09/14/2026
Input
1.4 per 1M tokens, unchanged since 09/14/2026
Output
4.4 per 1M tokens, unchanged since 09/14/2026

Last updated 10/03/2026, 02:10