prices synced 2026-09-11
Z

GLM 4.6V

Z.ai · Multimodal · Released Dec 2025

Compare

GLM 4.6V is Z.ai's multimodal model supporting text, image, and video inputs with a 131K-token context window.

Strengths
Handles multimodal reasoning across text, images, and video in a single request.
Best for
Tasks requiring analysis or reasoning over mixed media — documents with embedded images, video transcription with context, or cross-modal question answering.
Limitations
Superseded by GLM 5V Turbo, which offers expanded capabilities and is part of Z.ai's current foundation model generation.

Input / 1M

$0.3

Output / 1M

$0.9

Cached input / 1M

$0.055

Context window

131K

Price history

Price per 1M tokens over timeOutput $0.9; Input $0.3 as of Jun 2026.$0$0.2$0.4$0.6$0.8$1.00Output on 11 Jun 2026: $0.9Out $0.9Input on 11 Jun 2026: $0.3In $0.3Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.3 $0.9 $0.055 Imported from OpenRouter openrouter.ai

More from Z.ai

Report a problem