GPT-3.5 Turbo (batch)
OpenAI · Released May 2023
OpenAI's lightweight language model optimized for chat and code, available as a batch-processing variant for asynchronous workloads.
- Strengths
- Fast inference and lower resource consumption compared to larger models, while maintaining solid performance on natural language and code understanding tasks.
- Best for
- Asynchronous batch processing jobs where latency is not a constraint, such as bulk content generation, code analysis, or offline analytics pipelines.
- Limitations
- Smaller context window of 16,385 tokens limits it to shorter documents, and it has been superseded by newer models like GPT-4o and GPT-4o-mini that offer expanded capabilities.
Input / 1M
$0.25
Output / 1M
$0.75
Cached input / 1M
—
Context window
16K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 6 Aug 2026 | $0.25 | $0.75 | — | Imported from OpenRouter | openrouter.ai |