prices synced 2026-09-11
S

Step 3.5 Flash

StepFun · Released Jan 2026

Compare

Step 3.5 Flash is a lightweight language model from StepFun with a 262K token context window, designed for fast inference on tasks that don't require the full capacity of larger models.

Strengths
It offers fast inference speeds and reasonable performance for standard language tasks while maintaining a long context window for handling larger documents and conversations.
Best for
Applications prioritizing latency and throughput over maximum reasoning capability, such as real-time chat, content summarization, and document processing within a single request.
Limitations
It is superseded by Step 3.7 Flash (released May 2026), which uses sparse Mixture of Experts architecture for improved performance; Step 3.5 Flash is the older generation model and has lower capability across reasoning and complex tasks.

Input / 1M

$0.1

Output / 1M

$0.3

Cached input / 1M

Context window

262K

Price history

Input (solid)Output (dashed)
Price per 1M tokens over timeOutput went from $0.3 to $0.3; Input went from $0.09 to $0.1 between Jun 2026 and Jun 2026.$0$0.1$0.2$0.3Output on 11 Jun 2026: $0.3Output on 30 Jun 2026: $0.3Out $0.3Input on 11 Jun 2026: $0.09Input on 30 Jun 2026: $0.1In $0.1Jun 2026Jun 2026

Price change

30d
in out
90d
in increased 11.1% out decreased 0.0%
1y
in increased 11.1% out decreased 0.0%
Since launch
in increased 11.1% out decreased 0.0%

Snapshots

Effective Input Output Cached in Note Source
30 Jun 2026 $0.1 $0.3 Imported from OpenRouter openrouter.ai
11 Jun 2026 $0.09 $0.3 $0.02 Imported from OpenRouter openrouter.ai

More from StepFun

Report a problem