Qwen3.5-Flash
Alibaba · Multimodal · Released Feb 2026
Qwen3.5-Flash is a lightweight text and vision model from Alibaba with a 1 million token context window, positioned as a faster alternative to the mid-tier Qwen3.5 Plus and Qwen3.6 Plus.
- Strengths
- Low latency and efficient inference across both text and vision tasks without sacrificing broad capability support.
- Best for
- Real-time applications, streaming workloads, and vision tasks where speed and cost-efficiency are priorities over maximum accuracy.
- Limitations
- Smaller effective capacity than the Plus-tier models means it trades some accuracy and reasoning depth for speed and efficiency.
Input / 1M
$0.065
Output / 1M
$0.26
Cached input / 1M
—
Context window
1M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.065 | $0.26 | — | Imported from OpenRouter | openrouter.ai |