prices synced 2026-07-28

Qwen3 VL 235B A22B Thinking

Alibaba · Multimodal · Released Sep 2025

Compare

A 235-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.

Strengths
Handles complex visual understanding and reasoning tasks across multimodal inputs through its extended thinking mode, which outputs structured reasoning traces for multi-step problem solving.
Best for
Visual reasoning and analysis tasks requiring step-by-step problem decomposition, such as detailed image interpretation, chart analysis, and video understanding with explanatory output.
Limitations
The 131k-token context window is narrower than the 262k window of the 30B A3B Instruct, and the model is positioned as a reasoning variant rather than the general-purpose instruction-tuned counterpart in the Qwen3 VL 235B A22B line.

Input / 1M

$0.26

Output / 1M

$2.60

Cached input / 1M

Context window

131K

Price history

Price per 1M tokens over timeOutput $2.60; Input $0.26 as of Jun 2026.$0$1.00$2.00Output on 11 Jun 2026: $2.60Out $2.60Input on 11 Jun 2026: $0.26In $0.26Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.26 $2.60 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem