Qwen3.6 Flash
Alibaba · Multimodal · Released Apr 2026
Qwen3.6 Flash is Alibaba's lightweight language model from the Qwen3.6 generation with a 1 million token context window.
- Strengths
- Delivers fast inference across a broad range of tasks while maintaining a large context window.
- Best for
- Applications prioritizing latency and throughput, such as real-time chat, content generation, and long-context retrieval.
- Limitations
- As a lightweight model, it trades inference speed for lower reasoning capability compared to Qwen3.6 Plus and larger tier models in the Qwen3.6 generation.
Input / 1M
$0.1875
Output / 1M
$1.12
Cached input / 1M
—
Context window
1M
Also billed
- Cache write
- $0.2344 / 1M
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in decreased 0.0% out decreased 0.0%
- 1y
- in decreased 0.0% out decreased 0.0%
- Since launch
- in decreased 0.0% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 29 Jun 2026 | $0.1875 | $1.12 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.1875 | $1.12 | — | Imported from OpenRouter | openrouter.ai |