prices synced 2026-09-21
Z

GLM 5.2 (batch)

Z.ai · Released Jun 2026

Compare

GLM 5.2 (batch) is Z.ai's reasoning model optimized for asynchronous batch processing with a 1M-token context window.

Strengths
Handles extended documents and complex reasoning tasks efficiently in batch workflows without requiring real-time latency.
Best for
Non-urgent batch jobs that process large documents or require extended context, where throughput matters more than immediate response time.
Limitations
Designed for asynchronous batch processing and not suitable for synchronous or real-time inference scenarios; superseded by GLM 5.3 (batch) released in August 2026.

Input / 1M

$0.7

Output / 1M

$2.20

Cached input / 1M

$0.07

Context window

1.05M

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $2.20 to $2.20; Input went from $0.7 to $0.7 between Aug 2026 and Sep 2026.$0$1.00$2.00$3.00$4.00$5.00Output on 6 Aug 2026: $2.20Output on 19 Aug 2026: $4.40Output on 8 Sep 2026: $2.20Out $2.20Input on 6 Aug 2026: $0.7Input on 19 Aug 2026: $1.40Input on 8 Sep 2026: $0.7In $0.7Aug 2026Sep 2026

Price change

30d
in decreased 50.0% out decreased 50.0%
90d
in decreased 0.0% out decreased 0.0%
1y
in decreased 0.0% out decreased 0.0%
Since launch
in decreased 0.0% out decreased 0.0%

Snapshots

Effective Input Output Cached in Note Source
8 Sep 2026 $0.7 $2.20 $0.07 Imported from OpenRouter openrouter.ai
19 Aug 2026 $1.40 $4.40 $0.26 Imported from OpenRouter openrouter.ai
6 Aug 2026 $0.7 $2.20 $0.13 Imported from OpenRouter openrouter.ai

More from Z.ai

Report a problem