prices synced 2026-09-11

Qwen3.5-Flash

Alibaba · Multimodal · Released Feb 2026

Compare

Qwen3.5-Flash is a lightweight language model from Alibaba's Qwen3.5 generation with a 1 million token context window.

Strengths
Offers a substantially larger context window than other Qwen3.5 models while maintaining a smaller footprint than the larger parameter variants.
Best for
Applications requiring long-context processing without the computational overhead of larger models, such as document analysis and multi-turn conversations over extended sequences.
Limitations
Lacks the extended thinking and code specialization of Qwen3 Max Thinking and Qwen3 Coder Next, and is superseded by Qwen3.6 Flash from the newer generation.

Input / 1M

$0.065

Output / 1M

$0.26

Cached input / 1M

Context window

1M

Price history

Price per 1M tokens over timeOutput $0.26; Input $0.065 as of Jun 2026.$0$0.1$0.2Output on 11 Jun 2026: $0.26Out $0.26Input on 11 Jun 2026: $0.065In $0.065Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.065 $0.26 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem