Hermes 4 405B

Hermes 4 405B is a high-performance instruct and reasoning model from Nous Research with 405 billion parameters, designed for complex reasoning tasks and agentic workflows. The model supports optional thinking, precise tool calls, and structured outputs. Trained for high steerability and reduced Refusal rates. Available as an Open Weights model under the Meta Llama Community License.

NousResearch Version 4 Commercial use permitted Dense 405 B (405 B active) 128 K Context 01/2025 $1 / $3 per 1M

  • Restricted Weights
  • Server
  • OpenRouter
  • Text
  • Instruction-Tuned
  • Interactive

Sovereign Risk: LOW Nous Research is a US-based company and subject to the CLOUD Act; however, the weights are publicly available and can be run locally, so no third-party API access is required.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
68.4
Routine
42
Reasoning
26.4

Rank #76

LLM Judge Avg
3.35
100 Coverage
Avg Task Duration
21.36
Interactive
Token Rate
31.06
Output Rate
P95 Latency
51.12
Top 5 %
Total Tokens
57800
Output Volume
Cost per 1K
$0.003
USD / 1K Requests
Benchmark Cost
$0.17
Total · 57800 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Hermes 4 405B Best model Ø All models
Code Quality 67.96
CLI Benchmark 80.34
Logical Reasoning 61.65
UX Writing 61.61
Documentation 71.02
Content Transform. 73.47
Cultural Intelligence 68.64
Synthesis Quality 60
Tool Execution 90
ToolUse Score 68.5
Benchmark Cost $0.17

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile