Hermes 4 70B is a 70-billion parameter instruction-tuned model from Nous with a 131K token context window.
- Strengths
- It balances inference speed with capability, making it suitable for tasks requiring both performance and responsiveness without requiring the largest model size.
- Best for
- Multi-turn conversations, agentic workflows, and instruction-following tasks where 70B parameters meet latency or resource constraints.
- Limitations
- For reasoning-intensive or highly complex tasks, the 405B variants offer significantly more capacity, and this model sits between the earlier Hermes 3 70B and the newer Hermes 4 405B in the Nous lineup.
Input / 1M
$0.13
Output / 1M
$0.4
Cached input / 1M
—
Context window
131K
Price history
Snapshots
| Effective | Input | Output | Cached in | Note | Source |
|---|---|---|---|---|---|
| 11 Jun 2026 | $0.13 | $0.4 | — | Imported from OpenRouter | openrouter.ai |