GPT-4o-mini (batch)
OpenAI · Multimodal · Released Jul 2024
GPT-4o-mini (batch) is OpenAI's lightweight multimodal model optimized for asynchronous batch processing, handling text and images with a 128K token context window.
- Strengths
- Offers fast inference and efficient token usage for general-purpose tasks, maintaining multimodal capabilities at a smaller scale than GPT-4o.
- Best for
- Batch workflows where latency is not a constraint, such as bulk document analysis, content generation, and image understanding at scale.
- Limitations
- Smaller model capacity means reduced reasoning depth and accuracy on complex tasks compared to GPT-4o; not suitable for real-time or streaming applications.
Input / 1M
$0.075
Output / 1M
$0.3
Cached input / 1M
$0.0375
Context window
128K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 6 Aug 2026 | $0.075 | $0.3 | $0.0375 | Imported from OpenRouter | openrouter.ai |