Qwen3 VL 8B Thinking
Alibaba · Multimodal · Released Oct 2025
An 8-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.
- Strengths
- Combines compact model size with visual reasoning through a thinking variant that extends inference time for harder tasks across multimodal inputs.
- Best for
- Reasoning-heavy visual understanding tasks where model size and latency matter, such as document analysis, spatial reasoning, or complex scene interpretation on edge or cost-sensitive deployments.
- Limitations
- 8-billion parameters limit raw capability compared to larger variants like Qwen3 VL 30B A3B Thinking or Qwen3 VL 235B A22B Thinking, and extended reasoning adds latency that may not suit real-time applications.
Input / 1M
$0.18
Output / 1M
$2.10
Cached input / 1M
—
Context window
131K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in increased 53.8% out increased 53.8%
- 1y
- in increased 53.8% out increased 53.8%
- Since launch
- in increased 53.8% out increased 53.8%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 29 Jul 2026 | $0.18 | $2.10 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.117 | $1.36 | — | Imported from OpenRouter | openrouter.ai |