Back to models
available

GLM5.2-Fast

wafer-ai/glm5-2-fast

Prices

Input

$3.00 per 1M tokens

low
Input (cache write)

$0.00 per 1M tokens

low
Input (cached read)

$0.50 per 1M tokens

low
Output

$10.25 per 1M tokens

low

Specifications

Provider
wafer-ai
Context window
1,048,576
Max output tokens
131,072
Open weights
yes
Released
06/13/2026
Retires
no date announced
First seen by this tracker

Price history

No price of this model has changed since we started tracking it.

Input
3 per 1M tokens, unchanged since 09/14/2026
Output
10.25 per 1M tokens, unchanged since 09/14/2026
Input (cached read)
0.5 per 1M tokens, unchanged since 09/14/2026
Input (cache write)
0 per 1M tokens, unchanged since 09/14/2026

Last updated 10/04/2026, 02:10