I
Compare
Ling 3.0 Flash VL
inclusionAI · Multimodal · Released Sep 2026
Ling 3.0 Flash VL is a multimodal instruct model from InclusionAI that combines efficient language processing with native visual perception capabilities.
- Strengths
- It handles both text and image inputs with the speed and efficiency characteristic of the Flash tier, built on a mixture-of-experts architecture with 5.5B active parameters.
- Best for
- Applications requiring fast multimodal reasoning across text and image data within a very large context window.
- Limitations
- It is newer and less proven than the text-only Ling 3.0 Flash, so performance on specialized domains like finance may lag behind domain-specific variants like Ling 3.0 Flash Fin.
Input / 1M
$0.06
Output / 1M
$0.18
Cached input / 1M
$0.012
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Sep 2026 | $0.06 | $0.18 | $0.012 | Imported from OpenRouter | openrouter.ai |