prices synced 2026-09-20
I

Ling-3.0-flash

inclusionAI · Released Jul 2026

Compare

Ling-3.0-flash is a fast, efficient instruct model from inclusionAI with a 262K-token context window.

Strengths
It delivers low-latency inference while maintaining broad capability across general tasks with its large context window.
Best for
High-volume applications where inference speed and efficiency matter, such as real-time chat or batch processing with long documents.
Limitations
As a lighter model in the 3.0 lineup, it trades raw reasoning capability for speed; complex analytical or specialized domain tasks may benefit from larger parameter models.

Input / 1M

$0.021

Output / 1M

$0.063

Cached input / 1M

$0.0042

Context window

262K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $0.22 to $0.063; Input went from $0.075 to $0.021 between Aug 2026 and Aug 2026.$0$0.05$0.1$0.15$0.2$0.25Output on 6 Aug 2026: $0.22Output on 7 Aug 2026: $0.063Out $0.063Input on 6 Aug 2026: $0.075Input on 7 Aug 2026: $0.021In $0.021Aug 2026Aug 2026

Price change

30d
in out
90d
in decreased 72.0% out decreased 71.4%
1y
in decreased 72.0% out decreased 71.4%
Since launch
in decreased 72.0% out decreased 71.4%

Snapshots

Effective Input Output Cached in Note Source
7 Aug 2026 $0.021 $0.063 $0.0042 Imported from OpenRouter openrouter.ai
6 Aug 2026 $0.075 $0.22 $0.015 Imported from OpenRouter openrouter.ai

More from inclusionAI

Report a problem