prices synced 2026-09-11
I

Ling-2.6-flash

inclusionAI · Released Apr 2026

Compare

Ling-2.6-flash is a fast, efficient instruct model from inclusionAI with a 262K-token context window.

Strengths
It balances inference speed and efficiency with reasonable quality on a wide range of general tasks.
Best for
Applications where latency and throughput matter more than peak accuracy, or where you need to handle long documents within a single request.
Limitations
It trades off capability depth compared to larger models like Ling-2.6-1T, and has been superseded by Ling-3.0-flash in the provider's more recent releases.

Input / 1M

$0.01

Output / 1M

$0.03

Cached input / 1M

$0.002

Context window

262K

Price history

Price per 1M tokens over timeOutput $0.03; Input $0.01 as of Jun 2026.$0$0.01$0.02$0.03Output on 11 Jun 2026: $0.03Out $0.03Input on 11 Jun 2026: $0.01In $0.01Jun 2026

Snapshots

Effective Input Output Cached in Note Source
11 Jun 2026 $0.01 $0.03 $0.002 Imported from OpenRouter openrouter.ai

More from inclusionAI

Report a problem