prices synced 2026-09-21

GPT-4o (batch)

OpenAI · Multimodal · Released May 2024

Compare

GPT-4o (batch) is OpenAI's multimodal large language model optimized for asynchronous batch processing, handling text and images with a 128K token context window.

Strengths
It handles both text and images, maintains a large 128K token context window, and processes requests through batching for lower latency when immediate responses aren't required.
Best for
Workloads where you can defer response time in exchange for efficiency, such as bulk text generation, document processing, or image analysis of multiple files at once.
Limitations
Batch processing introduces latency compared to synchronous API calls, and more capable reasoning tasks may benefit from o1 (batch), which is trained with reinforcement learning for extended chain-of-thought processing.

Input / 1M

$1.25

Output / 1M

$5.00

Cached input / 1M

$0.625

Context window

128K

Price history

Price per 1M tokens over timeOutput $5.00; Input $1.25 as of Aug 2026.$0$1.00$2.00$3.00$4.00$5.00Output on 6 Aug 2026: $5.00Out $5.00Input on 6 Aug 2026: $1.25In $1.25Aug 2026

Snapshots

Effective Input Output Cached in Note Source
6 Aug 2026 $1.25 $5.00 $0.625 Imported from OpenRouter openrouter.ai

More from OpenAI

Report a problem