Step 3.5 Flash is a lightweight language model from StepFun with a 262K token context window, designed for fast inference on tasks that don't require the full capacity of larger models.
- Strengths
- It offers fast inference speeds and reasonable performance for standard language tasks while maintaining a long context window for handling larger documents and conversations.
- Best for
- Applications prioritizing latency and throughput over maximum reasoning capability, such as real-time chat, content summarization, and document processing within a single request.
- Limitations
- It is superseded by Step 3.7 Flash (released May 2026), which uses sparse Mixture of Experts architecture for improved performance; Step 3.5 Flash is the older generation model and has lower capability across reasoning and complex tasks.
Input / 1M
$0.1
Output / 1M
$0.3
Cached input / 1M
—
Context window
262K
Price history
Input (solid)Output (dashed)
Price change
- 30d
- in — out —
- 90d
- in increased 11.1% out decreased 0.0%
- 1y
- in increased 11.1% out decreased 0.0%
- Since launch
- in increased 11.1% out decreased 0.0%
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 30 Jun 2026 | $0.1 | $0.3 | — | Imported from OpenRouter | openrouter.ai |
| 11 Jun 2026 | $0.09 | $0.3 | $0.02 | Imported from OpenRouter | openrouter.ai |