Llama 8B (Unsloth, provenance unverified)
Dieses 8B-GGUF-Artefakt trägt einen Llama-3.3-Namen, aber Metas offizielles Llama-3.3 ist das 70B-Text-only-Modell — die hier gelistete 8B-Version ist eine inoffizielle Community-Konstruktion ohne verifizierte Upstream-Quelle. 8 Mrd. Dense-Parameter, 128.000 Tokens Kontext, lokal als Unsloth-GGUF betreibbar unter Llama-Community-Lizenz. Upstream-Herkunft bleibt bis zur Bestätigung unbestätigt.
- Restricted Weights
- Edge
- llama.cpp
- Text
- Instruction-Tuned
- Restricted-Weights
- Real-Time
Sovereign Risk: MEDIUM TODO
Schlüsselmetriken
Score · Latenz · Kosten · Qualität
- Total Score Bronze
- 62.71
- Routine
- 39.38
- Reasoning
- 23.33
- LLM Judge Avg
- 3.02 / 5
- 100 Coverage
- Avg Task Duration
- 15.17s
- Real-Time
- Token Rate
- 25.46tok/s
- Output Rate
- P95 Latency
- 31.25s
- Top 5 %
- Total Tokens
- 37200
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 37200 tok
Benchmark-Module
10 Module · gewichtet · vs. Modellmedian & Spitzenreiter
Llama 8B (Unsloth, provenance unverified)
Bestes Modell
Ø Alle Modelle
Code Quality
54.5
CLI Benchmark
82.22
Logical Reasoning
56.45
UX Writing
59.25
Documentation
54.68
Content Transform.
63.32
Cultural Intelligence
78.3
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost
$0
Token-Effizienz & Latenz
Verbrauch pro Modul vs. Modellmedian