Llama 8B (Unsloth, provenance unverified)
This 8B GGUF artifact carries a Llama-3.3 name, but Meta’s official Llama-3.3 is the 70B text-only model — the 8B version listed here is an unofficial community construction with no verified upstream source. 8B dense parameters, 128,000-token context, locally runnable as an Unsloth GGUF under the Llama Community License. Upstream provenance remains unconfirmed pending verification.
- Restricted Weights
- Edge
- llama.cpp
- Text
- Instruction-Tuned
- Restricted-Weights
- Real-Time
Sovereign Risk: MEDIUM TODO
Key metrics
Score · Latency · Cost · Quality
- Total Score Bronze
- 62.71
- Routine
- 39.38
- Reasoning
- 23.33
- LLM Judge Avg
- 3.02 / 5
- 100 Coverage
- Avg Task Duration
- 15.17s
- Real-Time
- Token Rate
- 25.46tok/s
- Output Rate
- P95 Latency
- 31.25s
- Top 5 %
- Total Tokens
- 37200
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 37200 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Llama 8B (Unsloth, provenance unverified)
Best model
Ø All models
Code Quality
54.5
CLI Benchmark
82.22
Logical Reasoning
56.45
UX Writing
59.25
Documentation
54.68
Content Transform.
63.32
Cultural Intelligence
78.3
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost
$0
Token efficiency & latency
Consumption per module vs. model median