Back to models
available

GLM-4.6V-Flash

z-ai/glm-4-6v-flash

Prices

Input

$0.00 per 1M tokens

low
Input (cache write)

$0.00 per 1M tokens

low
Input (cached read)

$0.00 per 1M tokens

low
Output

$0.00 per 1M tokens

low

Also available on

The same model, sold by someone else. The price above is the maker's own.

4 sales channels for this model.
ChannelWhatPrice
Other · huggingfaceInput$0.30 per 1M tokens
Output$0.90 per 1M tokens
Other · huggingface-novitaInput$0.30 per 1M tokens
Output$0.90 per 1M tokens
Other · temprInput$0.00 per 1M tokens
Input (cache write)$0.00 per 1M tokens
Input (cached read)$0.00 per 1M tokens
Output$0.00 per 1M tokens
Other · zenmuxInput$0.0218 per 1M tokens
Input (>32k ctx)$0.0437 per 1M tokens
Input (cached read, >32k ctx)$0.0044 per 1M tokens
Input (cached read)$0.0044 per 1M tokens
Output$0.2184 per 1M tokens
Output (>32k ctx)$0.4367 per 1M tokens

Specifications

Provider
z-ai
Context window
128,000
Max output tokens
32,768
Open weights
yes
Released
12/08/2025
Retires
no date announced
First seen by this tracker

Price history

No price of this model has changed since we started tracking it.

Input
0 per 1M tokens, unchanged since 09/22/2026
Output
0 per 1M tokens, unchanged since 09/22/2026
Input (cached read)
0 per 1M tokens, unchanged since 09/22/2026
Input (cache write)
0 per 1M tokens, unchanged since 09/22/2026

Last updated 10/03/2026, 02:10