I
Compare
Ling-2.6-flash
inclusionAI · Released Apr 2026
Ling-2.6-flash is a fast, efficient instruct model from inclusionAI with a 262K-token context window.
- Strengths
- It balances inference speed and efficiency with reasonable quality on a wide range of general tasks.
- Best for
- Applications where latency and throughput matter more than peak accuracy, or where you need to handle long documents within a single request.
- Limitations
- It trades off capability depth compared to larger models like Ling-2.6-1T, and has been superseded by Ling-3.0-flash in the provider's more recent releases.
Input / 1M
$0.01
Output / 1M
$0.03
Cached input / 1M
$0.002
Context window
262K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.01 | $0.03 | $0.002 | Imported from OpenRouter | openrouter.ai |