prices synced 2026-08-06

Compare two models

Select any two models. Input, output, cached, and context rarely move together — the cheaper model on one line isn't always the cheaper one on the next.

Description
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model from inclusionAI with approximately 5.1B parameters activated per token.
OpenAI's frontier model for complex professional workloads. 1M token context, text + image input, native computer use.
I/O price /1M
$0.075/$0.22
$5.00/$30.00
Input /1M
$0.075
$5.00
Output /1M
$0.22
$30.00
Cached in /1M
$0.015
$0.5
Context
131K
1M
Δ input since launch
Δ output since launch
Released
Jul 2026
Apr 2026
Status
Provider
I inclusionAI
OpenAI