prices synced 2026-09-11

Qwen3 VL 8B Thinking

Alibaba · Multimodal · Released Oct 2025

Compare

An 8-billion-parameter multimodal vision-language model from Alibaba that processes text, images, and video with extended reasoning capabilities and a 131k-token context window.

Strengths
Combines compact model size with visual reasoning through a thinking variant that extends inference time for harder tasks across multimodal inputs.
Best for
Reasoning-heavy visual understanding tasks where model size and latency matter, such as document analysis, spatial reasoning, or complex scene interpretation on edge or cost-sensitive deployments.
Limitations
8-billion parameters limit raw capability compared to larger variants like Qwen3 VL 30B A3B Thinking or Qwen3 VL 235B A22B Thinking, and extended reasoning adds latency that may not suit real-time applications.

Input / 1M

$0.18

Output / 1M

$2.10

Cached input / 1M

Context window

131K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $1.36 to $2.10; Input went from $0.117 to $0.18 between Jun 2026 and Jul 2026.$0$0.5$1.00$1.50$2.00Output on 11 Jun 2026: $1.36Output on 29 Jul 2026: $2.10Out $2.10Input on 11 Jun 2026: $0.117Input on 29 Jul 2026: $0.18In $0.18Jun 2026Jul 2026

Price change

30d
in out
90d
in increased 53.8% out increased 53.8%
1y
in increased 53.8% out increased 53.8%
Since launch
in increased 53.8% out increased 53.8%

Snapshots

Effective Input Output Cached in Note Source
29 Jul 2026 $0.18 $2.10 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.117 $1.36 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem