prices synced 2026-07-28

Qwen3 VL 30B A3B Thinking

Alibaba · Multimodal · Released Oct 2025

Compare

A 30B multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.

Strengths
Combines multimodal understanding across text, images, and video with reasoning chains for structured problem-solving across domain tasks.
Best for
Visual reasoning tasks that benefit from step-by-step inference—image analysis requiring explanation, video understanding with interpretation, or multimodal problems needing deliberate reasoning.
Limitations
The 30B parameter size sits between smaller and larger models in the Qwen3 VL lineup; reasoning capabilities are specialized and may add latency for tasks needing only fast inference without intermediate reasoning steps.

Input / 1M

$0.13

Output / 1M

$1.56

Cached input / 1M

Context window

131K

Price history

Price per 1M tokens over timeOutput $1.56; Input $0.13 as of Jun 2026.$0$0.5$1.00$1.50Output on 11 Jun 2026: $1.56Out $1.56Input on 11 Jun 2026: $0.13In $0.13Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.13 $1.56 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem