Back to models
available

GLM 4.6 Turbo (Thinking)

z-ai/glm-4-6-turbo:thinking

Prices

Input

$1.00 per 1M tokens

low
Input (cached read)

$0.50 per 1M tokens

low
Output

$3.00 per 1M tokens

low

Specifications

Provider
z-ai
Context window
204,800
Max output tokens
131,072
Open weights
yes
Released
10/02/2025
Retires
no date announced
First seen by this tracker

Price history

No price of this model has changed since we started tracking it.

Input
1 per 1M tokens, unchanged since 09/14/2026
Output
3 per 1M tokens, unchanged since 09/14/2026
Input (cached read)
0.5 per 1M tokens, unchanged since 09/14/2026

Last updated 10/02/2026, 02:10