Qwen3 VL 30B A3B Thinking
Alibaba · Multimodal · Released Oct 2025
A 30-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.
- Strengths
- Combines reasoning depth with moderate parameter count, handling complex visual and textual analysis across images and video without the computational demands of larger Thinking variants.
- Best for
- Vision-heavy tasks requiring step-by-step reasoning over images and video content where the 30B scale strikes a balance between capability and resource efficiency compared to the 235B Thinking model.
- Limitations
- Smaller than Qwen3 VL 235B A22B Thinking, so it trades raw reasoning capacity for efficiency; may struggle with tasks requiring the deepest levels of analysis that larger variants handle.
Input / 1M
$0.2
Output / 1M
$2.40
Cached input / 1M
—
Context window
131K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in increased 53.8% out increased 53.8%
- 1y
- in increased 53.8% out increased 53.8%
- Since launch
- in increased 53.8% out increased 53.8%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 29 Jul 2026 | $0.2 | $2.40 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.13 | $1.56 | — | Imported from OpenRouter | openrouter.ai |