Hermes 4 70B

Hermes 4 70B is an open instruct and reasoning model by Nous Research from the Hermes 4 family with 70 billion parameters. The model combines optional thinking with advanced tool use and structured outputs, trained for high steerability and reduced Refusal rates. Available as an Open Weights model under a Modified MIT license for local or server-side deployment.

NousResearch Version 4 Commercial use permitted Dense 70 B (70 B active) 131 K Context 01/2025 $0.13 / $0.4 per 1M

  • Open Weights
  • Server
  • OpenRouter
  • Text
  • Instruction-Tuned
  • Real-Time

Sovereign Risk: MEDIUM Nous Research is a US-based company, which makes US jurisdiction and the CLOUD Act relevant. Since the models are Open Weights, they can be run locally, which minimizes the risk of uncontrolled data leakage via an API.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
68.99
Routine
41.2
Reasoning
27.79

Rank #73

LLM Judge Avg
3.47
100 Coverage
Avg Task Duration
10.05
Real-Time
Token Rate
76.83
Output Rate
P95 Latency
20.54
Top 5 %
Total Tokens
64099.99999999999
Output Volume
Cost per 1K
$0.0004
USD / 1K Requests
Benchmark Cost
$0.03
Total · 64099.99999999999 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Hermes 4 70B Best model Ø All models
Code Quality 65.24
CLI Benchmark 78
Logical Reasoning 73
UX Writing 62.08
Documentation 65.14
Content Transform. 63.06
Cultural Intelligence 74.4
Synthesis Quality 41.67
Tool Execution 88.33
ToolUse Score 75.5
Benchmark Cost $0.03

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile