GPT-4o (2024-05-13)
OpenAI · Multimodal · Released May 2024
A multimodal large language model from OpenAI that processes text and images with a 128K token context window.
- Strengths
- Handles both text and image inputs natively, balancing capability and inference speed across a wide range of tasks.
- Best for
- General-purpose applications that need multimodal reasoning without the extended thinking overhead of later reasoning models.
- Limitations
- Superseded by GPT-4o (2024-11-20) in November 2024, which added native structured output support; for real-time information, use GPT-4o Search Preview instead.
Input / 1M
$5.00
Output / 1M
$15.00
Cached input / 1M
—
Context window
128K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $5.00 | $15.00 | — | Imported from OpenRouter | openrouter.ai |