Hermes 4 70B
Hermes 4 70B is an open instruct and reasoning model by Nous Research from the Hermes 4 family with 70 billion parameters. The model combines optional thinking with advanced tool use and structured outputs, trained for high steerability and reduced Refusal rates. Available as an Open Weights model under a Modified MIT license for local or server-side deployment.
- Open Weights
- Server
- OR
- Text
- Instruction-Tuned
- Real-Time
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 70.44
- Routine
- 43.88
- Reasoning
- 26.56
- LLM Judge Avg
- 3.56 / 5
- 100 Coverage
- Avg Task Duration
- 7.38s
- Real-Time
- Token Rate
- 75.8tok/s
- Output Rate
- P95 Latency
- 24.38s
- Top 5 %
- Total Tokens
- 52600
- Output Volume
- Cost per 1K
- $0.0004
- USD / 1K Requests
- Benchmark Cost
- $0.02
- Total · 52600 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median