Back to models
available

GLM 4.1V Thinking Flash

nano-gpt/glm-4-1v-thinking-flash

Prices

Input

$0.30 per 1M tokens

low
Input (cached read)

$0.15 per 1M tokens

low
Output

$0.30 per 1M tokens

low

Specifications

Provider
nano-gpt
Context window
64,000
Max output tokens
8,192
Open weights
no
Released
07/09/2025
Retires
no date announced
First seen by this tracker

Price history

No price of this model has changed since we started tracking it.

Input
0.3 per 1M tokens, unchanged since 09/14/2026
Output
0.3 per 1M tokens, unchanged since 09/14/2026
Input (cached read)
0.15 per 1M tokens, unchanged since 09/14/2026

Last updated 10/03/2026, 02:10