prices synced 2026-07-28

Compare two models

Select any two models. Input, output, cached, and context rarely move together — the cheaper model on one line isn't always the cheaper one on the next.

Description
Ling-2.6-flash is a 104B parameter model from inclusionAI with 7.4B active parameters, designed for fast inference and efficient token usage in agent applications.
OpenAI's frontier model for complex professional workloads. 1M token context, text + image input, native computer use.
I/O price /1M
$0.01/$0.03
$5.00/$30.00
Input /1M
$0.01
$5.00
Output /1M
$0.03
$30.00
Cached in /1M
$0.002
$0.5
Context
262K
1M
Δ input since launch
Δ output since launch
Released
Apr 2026
Apr 2026
Status
Provider
I inclusionAI
OpenAI