GLM 5.2 (batch) is Z.ai's reasoning model optimized for asynchronous batch processing with a 1M-token context window.
- Strengths
- Handles extended documents and complex reasoning tasks efficiently in batch workflows without requiring real-time latency.
- Best for
- Non-urgent batch jobs that process large documents or require extended context, where throughput matters more than immediate response time.
- Limitations
- Designed for asynchronous batch processing and not suitable for synchronous or real-time inference scenarios; superseded by GLM 5.3 (batch) released in August 2026.
Input / 1M
$0.7
Output / 1M
$2.20
Cached input / 1M
$0.07
Context window
1.05M
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in decreased 50.0% out decreased 50.0%
- 90d
- in decreased 0.0% out decreased 0.0%
- 1y
- in decreased 0.0% out decreased 0.0%
- Since launch
- in decreased 0.0% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 8 Sep 2026 | $0.7 | $2.20 | $0.07 | Imported from OpenRouter | openrouter.ai |
| 19 Aug 2026 | $1.40 | $4.40 | $0.26 | Imported from OpenRouter | openrouter.ai |
| 6 Aug 2026 | $0.7 | $2.20 | $0.13 | Imported from OpenRouter | openrouter.ai |