Llama 8B (Unsloth, provenance unverified)

This 8B GGUF artifact carries a Llama-3.3 name, but Meta’s official Llama-3.3 is the 70B text-only model — the 8B version listed here is an unofficial community construction with no verified upstream source. 8B dense parameters, 128,000-token context, locally runnable as an Unsloth GGUF under the Llama Community License. Upstream provenance remains unconfirmed pending verification.

Meta Version 3 Commercial use permitted Dense 8 B (8 B active) 128 K Context 12/2023 locally tested

  • Restricted Weights
  • Edge
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Restricted-Weights
  • Real-Time

Sovereign Risk: MEDIUM TODO

Key metrics

Score · Latency · Cost · Quality

Total Score Bronze
62.71
Routine
39.38
Reasoning
23.33

Rank #90

LLM Judge Avg
3.02
100 Coverage
Avg Task Duration
15.17
Real-Time
Token Rate
25.46
Output Rate
P95 Latency
31.25
Top 5 %
Total Tokens
37200
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 37200 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Llama 8B (Unsloth, provenance unverified) Best model Ø All models
Code Quality 54.5
CLI Benchmark 82.22
Logical Reasoning 56.45
UX Writing 59.25
Documentation 54.68
Content Transform. 63.32
Cultural Intelligence 78.3
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile