GPT-3.5 Turbo (older v0613)
OpenAI · Released Jan 2024
GPT-3.5 Turbo v0613 is an older snapshot of OpenAI's lightweight language model, optimized for chat and code tasks with a 4K token context window.
- Strengths
- Fast inference and low resource requirements make it suitable for latency-sensitive applications and high-volume workloads.
- Best for
- Simple text generation, chat-like interactions, and code completion where speed is prioritized over reasoning depth.
- Limitations
- The 4K token context window severely limits document processing and multi-turn conversation history; this version is outdated and superseded by newer GPT-3.5 Turbo releases and the multimodal GPT-4o family, and lacks support for images and structured output.
Input / 1M
$1.00
Output / 1M
$2.00
Cached input / 1M
—
Context window
4K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $1.00 | $2.00 | — | Imported from OpenRouter | openrouter.ai |