Qwen3 VL 235B A22B Instruct
Alibaba · Multimodal · Released Sep 2025
A 235-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with a 131k-token context window.
- Strengths
- Handles complex multimodal tasks across long documents and conversations through its large parameter count and extensive context window.
- Best for
- Long-context vision-language tasks requiring deep reasoning over documents with images and video, where model scale and capacity matter.
- Limitations
- The 131k-token context is narrower than some alternatives like Qwen3 Next 80B A3B Instruct (262k tokens), and lacks the extended reasoning capabilities of its sibling Qwen3 VL 235B A22B Thinking.
Input / 1M
$0.21
Output / 1M
$1.90
Cached input / 1M
$0.1
Context window
131K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in decreased 19.2% out increased 82.7%
- 90d
- in increased 5.0% out increased 115.9%
- 1y
- in increased 5.0% out increased 115.9%
- Since launch
- in increased 5.0% out increased 115.9%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 17 Aug 2026 | $0.21 | $1.90 | $0.1 | Imported from OpenRouter | openrouter.ai |
| 12 Aug 2026 | $0.26 | $1.04 | — | Imported from OpenRouter | openrouter.ai |
| 15 Jul 2026 | $0.21 | $1.90 | $0.1 | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.2 | $0.88 | $0.11 | Imported from OpenRouter | openrouter.ai |