prices synced 2026-09-11

Qwen3.8 Flash

Alibaba · Multimodal · Released Aug 2026

Compare

Qwen3.8 Flash is Alibaba's lightweight multimodal reasoning model supporting text, image, and video inputs with a 1 million token context window.

Strengths
Handles coding assistance, visual understanding, document analysis, and long-video processing while maintaining fast inference throughput.
Best for
Agentic workflows, desktop interaction, chart analysis, and codebase reasoning where multimodal input and reasoning speed matter.
Limitations
A faster tier compared to Qwen3.8 Max; trade-offs in reasoning depth and parameter count are typical for models optimized for latency.

Input / 1M

$0.15

Output / 1M

$0.47

Cached input / 1M

$0.016

Context window

1M

Also billed

Cache write
$0.2 / 1M

Price history

Price per 1M tokens over timeOutput $0.47; Input $0.15 as of Aug 2026.$0$0.1$0.2$0.3$0.4$0.5Output on 27 Aug 2026: $0.47Out $0.47Input on 27 Aug 2026: $0.15In $0.15Aug 2026

Snapshots

Effective Input Output Cached in Note Source
27 Aug 2026 $0.15 $0.47 $0.016 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem