GLM 4.7 is Z.ai's foundation model with a 202K-token context window for processing extended documents and conversations.
- Strengths
- GLM 4.7 handles long-context tasks effectively with a 202K-token window, positioning it as a capable general-purpose foundation model across a range of text-based workloads.
- Best for
- Processing extended documents, multi-turn conversations, and tasks requiring broad language understanding without specialized needs for speed optimization or multimodal input.
- Limitations
- GLM 4.7 is superseded by newer Z.ai models—GLM 5 (released 2026-02) and beyond offer improved capabilities, while GLM 4.7 Flash (released 2026-01) is the optimized inference variant if latency is a concern.
Input / 1M
$0.4
Output / 1M
$1.75
Cached input / 1M
$0.08
Context window
202K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.4 | $1.75 | $0.08 | Imported from OpenRouter | openrouter.ai |