prices synced 2026-09-11
I

Ling 3.0 Flash VL

inclusionAI · Multimodal · Released Sep 2026

Compare

Ling 3.0 Flash VL is a multimodal instruct model from InclusionAI that combines efficient language processing with native visual perception capabilities.

Strengths
It handles both text and image inputs with the speed and efficiency characteristic of the Flash tier, built on a mixture-of-experts architecture with 5.5B active parameters.
Best for
Applications requiring fast multimodal reasoning across text and image data within a very large context window.
Limitations
It is newer and less proven than the text-only Ling 3.0 Flash, so performance on specialized domains like finance may lag behind domain-specific variants like Ling 3.0 Flash Fin.

Input / 1M

$0.06

Output / 1M

$0.18

Cached input / 1M

$0.012

Context window

131K

Price history

Price per 1M tokens over timeOutput $0.18; Input $0.06 as of Sep 2026.$0$0.05$0.1$0.15$0.2Output on 11 Sep 2026: $0.18Out $0.18Input on 11 Sep 2026: $0.06In $0.06Sep 2026

Snapshots

Effective Input Output Cached in Note Source
11 Sep 2026 $0.06 $0.18 $0.012 Imported from OpenRouter openrouter.ai

More from inclusionAI

Report a problem