prices synced 2026-07-28

Llama 4 Scout

Meta · Released Apr 2025

Compare

Meta's smaller open-weight MoE model (17B active / 109B total). Supports up to 10M context on some providers.

Strengths
Smaller open MoE supporting up to a 10M-token context on some hosts.
Best for
Extreme long-context tasks and budget self-hosting.
Limitations
Lighter than Maverick; hosted pricing and context limits vary by provider.

Input / 1M

$0.08

Output / 1M

$0.3

Cached input / 1M

Context window

1M

Price history

Price per 1M tokens over timeOutput $0.3; Input $0.08 as of Apr 2025.$0$0.1$0.2$0.3Output on 5 Apr 2025: $0.3Out $0.3Input on 5 Apr 2025: $0.08In $0.08Apr 2025

Snapshots

Effective Input Output Cached in Note Source
5 Apr 2025 $0.08 $0.3 Representative hosted rate (DeepInfra); varies by provider pricepertoken.com

More from Meta

Report a problem