prices synced 2026-07-28
Z

GLM 5 Turbo

Z.ai · Released Mar 2026

Compare

GLM 5 Turbo is Z.ai's optimized inference variant of the GLM 5 foundation model, designed for faster response times while maintaining broad capability across reasoning and long-context tasks.

Strengths
It handles complex reasoning, code, and long documents efficiently within its 262K-token context window, balancing speed and quality across general workloads.
Best for
Applications requiring solid general-purpose reasoning and long-context processing where faster inference is preferred over maximum model scale, such as real-time agent systems and document analysis.
Limitations
It is positioned between the base GLM 5 and the reasoning-optimized GLM 5.2; workloads demanding maximum reasoning depth or the 1M-token context of GLM 5.2 may benefit from those models instead.

Input / 1M

$1.20

Output / 1M

$4.00

Cached input / 1M

$0.24

Context window

202K

Price history

Price per 1M tokens over timeOutput $4.00; Input $1.20 as of Jun 2026.$0$1.00$2.00$3.00$4.00Output on 11 Jun 2026: $4.00Out $4.00Input on 11 Jun 2026: $1.20In $1.20Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $1.20 $4.00 $0.24 Imported from OpenRouter openrouter.ai

More from Z.ai

Report a problem