prices synced 2026-09-11

Qwen3 VL 30B A3B Thinking

Alibaba · Multimodal · Released Oct 2025

Compare

A 30-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.

Strengths
Combines reasoning depth with moderate parameter count, handling complex visual and textual analysis across images and video without the computational demands of larger Thinking variants.
Best for
Vision-heavy tasks requiring step-by-step reasoning over images and video content where the 30B scale strikes a balance between capability and resource efficiency compared to the 235B Thinking model.
Limitations
Smaller than Qwen3 VL 235B A22B Thinking, so it trades raw reasoning capacity for efficiency; may struggle with tasks requiring the deepest levels of analysis that larger variants handle.

Input / 1M

$0.2

Output / 1M

$2.40

Cached input / 1M

Context window

131K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $1.56 to $2.40; Input went from $0.13 to $0.2 between Jun 2026 and Jul 2026.$0$0.5$1.00$1.50$2.00$2.50Output on 11 Jun 2026: $1.56Output on 29 Jul 2026: $2.40Out $2.40Input on 11 Jun 2026: $0.13Input on 29 Jul 2026: $0.2In $0.2Jun 2026Jul 2026

Price change

30d
in out
90d
in increased 53.8% out increased 53.8%
1y
in increased 53.8% out increased 53.8%
Since launch
in increased 53.8% out increased 53.8%

Snapshots

Effective Input Output Cached in Note Source
29 Jul 2026 $0.2 $2.40 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.13 $1.56 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem