prices synced 2026-09-11

Qwen3.7 Flash

Alibaba · Multimodal · Released Jul 2026

Compare

Qwen3.7 Flash is Alibaba's lightweight language model from the Qwen3.7 generation with a 1 million token context window.

Strengths
It prioritizes speed and efficiency while maintaining a very large context window for handling long documents and conversations.
Best for
Applications that need fast inference on text workloads with the ability to process large inputs.
Limitations
As a lightweight model in the Qwen3.7 lineup, it trades capability depth for speed compared to Qwen3.7 Plus and Qwen3.7 Max.

Input / 1M

$0.03

Output / 1M

$0.13

Cached input / 1M

$0.006

Context window

1M

Also billed

Cache write
$0.038 / 1M

Price history

Price per 1M tokens over timeOutput $0.13; Input $0.03 as of Jul 2026.$0$0.05$0.1Output on 28 Jul 2026: $0.13Out $0.13Input on 28 Jul 2026: $0.03In $0.03Jul 2026

Snapshots

Effective Input Output Cached in Note Source
28 Jul 2026 $0.03 $0.13 $0.006 Imported from OpenRouter openrouter.ai

More from Alibaba

Report a problem