Compare two models
Select any two models. Input, output, cached, and context rarely move together — the cheaper model on one line isn't always the cheaper one on the next.
Description
Ling-2.6-flash is a 104B parameter model from inclusionAI with 7.4B active parameters, designed for fast inference and efficient token usage in agent applications.
OpenAI's frontier model for complex professional workloads. 1M token context, text + image input, native computer use.
I/O price /1M
$0.01/$0.03
$5.00/$30.00
Input /1M
$0.01
$5.00
Output /1M
$0.03
$30.00
Cached in /1M
$0.002
$0.5
Context
262K
1M
Δ input since launch
—
—
Δ output since launch
—
—
Released
Apr 2026
Apr 2026
Status
—
—
Provider
I
inclusionAI
OpenAI