prices synced 2026-08-07

GPT-3.5 Turbo (batch)

OpenAI · Released May 2023

Compare

OpenAI's lightweight language model optimized for chat and code, available as a batch-processing variant for asynchronous workloads.

Strengths
Fast inference and lower resource consumption compared to larger models, while maintaining solid performance on natural language and code understanding tasks.
Best for
Asynchronous batch processing jobs where latency is not a constraint, such as bulk content generation, code analysis, or offline analytics pipelines.
Limitations
Smaller context window of 16,385 tokens limits it to shorter documents, and it has been superseded by newer models like GPT-4o and GPT-4o-mini that offer expanded capabilities.

Input / 1M

$0.25

Output / 1M

$0.75

Cached input / 1M

Context window

16K

Price history

Price per 1M tokens over timeOutput $0.75; Input $0.25 as of Aug 2026.$0$0.2$0.4$0.6$0.8Output on 6 Aug 2026: $0.75Out $0.75Input on 6 Aug 2026: $0.25In $0.25Aug 2026

Snapshots

Effective Input Output Cached in Note Source
6 Aug 2026 $0.25 $0.75 Imported from OpenRouter openrouter.ai

More from OpenAI

Report a problem