Hermes 4 405B
Hermes 4 405B is a high-performance instruct and reasoning model from Nous Research with 405 billion parameters, designed for complex reasoning tasks and agentic workflows. The model supports optional thinking, precise tool calls, and structured outputs. Trained for high steerability and reduced Refusal rates. Available as an Open Weights model under the Meta Llama Community License.
- Restricted Weights
- Frontier
- OR
- Text
- Instruction-Tuned
- Real-Time
Sovereign Risk: LOW Nous Research is a US-based company and subject to the CLOUD Act; however, the weights are publicly available and can be run locally, so no third-party API access is required.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 67.76
- Routine
- 40.87
- Reasoning
- 26.9
- LLM Judge Avg
- 3.3 / 5
- 100 Coverage
- Avg Task Duration
- 18.37s
- Real-Time
- Token Rate
- 39.49tok/s
- Output Rate
- P95 Latency
- 37.76s
- Top 5 %
- Total Tokens
- 48900
- Output Volume
- Cost per 1K
- $0.003
- USD / 1K Requests
- Benchmark Cost
- $0.15
- Total · 48900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median