GPT-4o (batch)
OpenAI · Multimodal · Released May 2024
GPT-4o (batch) is OpenAI's multimodal large language model optimized for asynchronous batch processing, handling text and images with a 128K token context window.
- Strengths
- It handles both text and images, maintains a large 128K token context window, and processes requests through batching for lower latency when immediate responses aren't required.
- Best for
- Workloads where you can defer response time in exchange for efficiency, such as bulk text generation, document processing, or image analysis of multiple files at once.
- Limitations
- Batch processing introduces latency compared to synchronous API calls, and more capable reasoning tasks may benefit from o1 (batch), which is trained with reinforcement learning for extended chain-of-thought processing.
Input / 1M
$1.25
Output / 1M
$5.00
Cached input / 1M
$0.625
Context window
128K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 6 Aug 2026 | $1.25 | $5.00 | $0.625 | Imported from OpenRouter | openrouter.ai |