Hermes 4 70B

Hermes 4 70B is an open instruct and reasoning model by Nous Research from the Hermes 4 family with 70 billion parameters. The model combines optional thinking with advanced tool use and structured outputs, trained for high steerability and reduced Refusal rates. Available as an Open Weights model under a Modified MIT license for local or server-side deployment.

NousResearch Version 4 Commercial use permitted Dense 70 B (70 B active) 131 K Context 01/2025 $0.13 / $0.4 per 1M

  • Open Weights
  • Server
  • OR
  • Text
  • Instruction-Tuned
  • Real-Time

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
70.44
Routine
43.88
Reasoning
26.56

Rank #63

LLM Judge Avg
3.56
100 Coverage
Avg Task Duration
7.38
Real-Time
Token Rate
75.8
Output Rate
P95 Latency
24.38
Top 5 %
Total Tokens
52600
Output Volume
Cost per 1K
$0.0004
USD / 1K Requests
Benchmark Cost
$0.02
Total · 52600 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Hermes 4 70B Best model Ø All models
Code Quality 62.96
CLI Benchmark 87.34
Logical Reasoning 66.42
UX Writing 68.67
Documentation 74.63
Content Transform. 74.17
Cultural Intelligence 68.92
Synthesis Quality 41.67
Tool Execution 88.33
ToolUse Score 66.42
Benchmark Cost $0.02

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile