Hermes 4 14B (Abliterated)

Hermes 4 14B Abliterated is a locally deployable Open Weights variant by NousResearch based on Qwen 3 14B, with safety mechanisms deliberately removed. With 14 billion parameters and a 128,000-token context window, the model targets unfiltered response behavior, creative writing workflows, and scenarios requiring maximum openness. Fully commercially usable under the Apache 2.0 license.

NousResearch Version 4.0 Commercial use permitted Dense 14 B (14 B active) 128 K Context 09/2024 locally tested

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Uncensored
  • Batch

Sovereign Risk: MEDIUM NousResearch is a US-based company; the CLOUD Act is only relevant when using the API, not when running the Open Weights variant locally. Abliteration only increases behavioral openness, not the provenance risk of the weights.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
68.19
Routine
41.26
Reasoning
26.92

Rank #75

LLM Judge Avg
3.27
100 Coverage
Avg Task Duration
65.68
Batch
Token Rate
25.39
Output Rate
P95 Latency
299.57
Top 5 %
Total Tokens
63400
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 63400 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Hermes 4 14B (Abliterated) Best model Ø All models
Code Quality 62.9
CLI Benchmark 85.56
Logical Reasoning 64.55
UX Writing 61.55
Documentation 64.63
Content Transform. 74.43
Cultural Intelligence 71
Synthesis Quality 36.67
Tool Execution 90
ToolUse Score 69.54
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile