Back to models
available
GLM-4.6V-Flash
z-ai/glm-4-6v-flashPrices
- Input
$0.00 per 1M tokens
low- Input (cache write)
$0.00 per 1M tokens
low- Input (cached read)
$0.00 per 1M tokens
low- Output
$0.00 per 1M tokens
low
Also available on
The same model, sold by someone else. The price above is the maker's own.
| Channel | What | Price |
|---|---|---|
| Other · huggingface | Input | $0.30 per 1M tokens |
| Output | $0.90 per 1M tokens | |
| Other · huggingface-novita | Input | $0.30 per 1M tokens |
| Output | $0.90 per 1M tokens | |
| Other · tempr | Input | $0.00 per 1M tokens |
| Input (cache write) | $0.00 per 1M tokens | |
| Input (cached read) | $0.00 per 1M tokens | |
| Output | $0.00 per 1M tokens | |
| Other · zenmux | Input | $0.0218 per 1M tokens |
| Input (>32k ctx) | $0.0437 per 1M tokens | |
| Input (cached read, >32k ctx) | $0.0044 per 1M tokens | |
| Input (cached read) | $0.0044 per 1M tokens | |
| Output | $0.2184 per 1M tokens | |
| Output (>32k ctx) | $0.4367 per 1M tokens |
Specifications
- Provider
- z-ai
- Context window
- 128,000
- Max output tokens
- 32,768
- Open weights
- yes
- Released
- 12/08/2025
- Retires
- no date announced
- First seen by this tracker
Price history
No price of this model has changed since we started tracking it.
- Input
- 0 per 1M tokens, unchanged since 09/22/2026
- Output
- 0 per 1M tokens, unchanged since 09/22/2026
- Input (cached read)
- 0 per 1M tokens, unchanged since 09/22/2026
- Input (cache write)
- 0 per 1M tokens, unchanged since 09/22/2026