Qwen3 VL 8B Thinking
Alibaba · Multimodal · Released Oct 2025
An 8B multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities for structured problem-solving.
- Strengths
- Combines visual understanding with step-by-step reasoning chains, enabling it to handle complex multimodal tasks that require explicit reasoning traces.
- Best for
- Visual reasoning tasks on constrained resources—where you need multimodal capabilities and transparent reasoning without the parameter count of larger thinking models.
- Limitations
- Smaller parameter count means it trades off raw reasoning depth compared to the 30B and 235B Thinking variants in the lineup, and visual complexity handling is limited relative to larger models.
Input / 1M
$0.117
Output / 1M
$1.36
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.117 | $1.36 | — | Imported from OpenRouter | openrouter.ai |