prices synced 2026-07-28
Z

GLM 4.6V

Z.ai · Multimodal · Released Dec 2025

Compare

GLM 4.6V is Z.ai's vision-language model that processes text, image, and video inputs for multimodal reasoning tasks.

Strengths
Handles multimodal inputs across images and video alongside text, enabling vision-grounded analysis and understanding.
Best for
Applications requiring simultaneous processing of visual and textual content, such as image understanding, video analysis, and cross-modal reasoning.
Limitations
Older than the current generation of Z.ai models; GLM 5V Turbo and GLM 5.2 supersede this model with newer architectures and larger context windows.

Input / 1M

$0.3

Output / 1M

$0.9

Cached input / 1M

$0.055

Context window

131K

Price history

Price per 1M tokens over timeOutput $0.9; Input $0.3 as of Jun 2026.$0$0.2$0.4$0.6$0.8$1.00Output on 11 Jun 2026: $0.9Out $0.9Input on 11 Jun 2026: $0.3In $0.3Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.3 $0.9 $0.055 Imported from OpenRouter openrouter.ai

More from Z.ai

Report a problem