Qwen3 VL 235B A22B Thinking
Alibaba · Multimodal · Released Sep 2025
A 235-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.
- Strengths
- Handles complex visual understanding and reasoning tasks across multimodal inputs through its extended thinking mode, which outputs structured reasoning traces for multi-step problem solving.
- Best for
- Visual reasoning and analysis tasks requiring step-by-step problem decomposition, such as detailed image interpretation, chart analysis, and video understanding with explanatory output.
- Limitations
- The 131k-token context window is narrower than the 262k window of the 30B A3B Instruct, and the model is positioned as a reasoning variant rather than the general-purpose instruction-tuned counterpart in the Qwen3 VL 235B A22B line.
Input / 1M
$0.26
Output / 1M
$2.60
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.26 | $2.60 | — | Imported from OpenRouter | openrouter.ai |