prices synced 2026-07-28

Qwen3.5-Flash

Alibaba · Multimodal · Released Feb 2026

Compare

Qwen3.5-Flash is a lightweight text and vision model from Alibaba with a 1 million token context window, positioned as a faster alternative to the mid-tier Qwen3.5 Plus and Qwen3.6 Plus.

Strengths
Low latency and efficient inference across both text and vision tasks without sacrificing broad capability support.
Best for
Real-time applications, streaming workloads, and vision tasks where speed and cost-efficiency are priorities over maximum accuracy.
Limitations
Smaller effective capacity than the Plus-tier models means it trades some accuracy and reasoning depth for speed and efficiency.

Input / 1M

$0.065

Output / 1M

$0.26

Cached input / 1M

Context window

1M

Price history

Price per 1M tokens over timeOutput $0.26; Input $0.065 as of Jun 2026.$0$0.1$0.2Output on 11 Jun 2026: $0.26Out $0.26Input on 11 Jun 2026: $0.065In $0.065Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.065 $0.26 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem