I
Compare
Ling-3.0-flash
inclusionAI · Released Jul 2026
Ling-3.0-flash is a fast, efficient instruct model from inclusionAI with a 262K-token context window.
- Strengths
- It delivers low-latency inference while maintaining broad capability across general tasks with its large context window.
- Best for
- High-volume applications where inference speed and efficiency matter, such as real-time chat or batch processing with long documents.
- Limitations
- As a lighter model in the 3.0 lineup, it trades raw reasoning capability for speed; complex analytical or specialized domain tasks may benefit from larger parameter models.
Input / 1M
$0.021
Output / 1M
$0.063
Cached input / 1M
$0.0042
Context window
262K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in decreased 72.0% out decreased 71.4%
- 1y
- in decreased 72.0% out decreased 71.4%
- Since launch
- in decreased 72.0% out decreased 71.4%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 7 Aug 2026 | $0.021 | $0.063 | $0.0042 | Imported from OpenRouter | openrouter.ai |
| 6 Aug 2026 | $0.075 | $0.22 | $0.015 | Imported from OpenRouter | openrouter.ai |