Back to models
available
GLM 4.1V Thinking Flash
nano-gpt/glm-4-1v-thinking-flashPrices
- Input
$0.30 per 1M tokens
low- Input (cached read)
$0.15 per 1M tokens
low- Output
$0.30 per 1M tokens
low
Specifications
- Provider
- nano-gpt
- Context window
- 64,000
- Max output tokens
- 8,192
- Open weights
- no
- Released
- 07/09/2025
- Retires
- no date announced
- First seen by this tracker
Price history
No price of this model has changed since we started tracking it.
- Input
- 0.3 per 1M tokens, unchanged since 09/14/2026
- Output
- 0.3 per 1M tokens, unchanged since 09/14/2026
- Input (cached read)
- 0.15 per 1M tokens, unchanged since 09/14/2026