GPT-4.1 Mini (batch)
OpenAI · Multimodal · Released Apr 2025
OpenAI's GPT-4.1 Mini optimized for batch processing, with 1 million token context window for asynchronous workloads.
- Strengths
- Runs fast and efficient inference on large document batches without token usage or latency constraints.
- Best for
- High-volume classification, extraction, and summarization tasks submitted as batch jobs over hours or days.
- Limitations
- Limited to batch API with asynchronous processing; slower real-time performance than the standard GPT-4.1 Mini, and does not support reasoning or chain-of-thought modes.
Input / 1M
$0.2
Output / 1M
$0.8
Cached input / 1M
$0.05
Context window
1.05M
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 6 Aug 2026 | $0.2 | $0.8 | $0.05 | Imported from OpenRouter | openrouter.ai |