Llama 3.2 3B (Unsloth)

3,21 Mrd. Dense-Parameter, 128.000 Tokens Kontext: Llama 3.2 3B ist Metas kompakte Text-only-Variante der Llama-3.2-Familie für lokale Aufgaben wie Zusammenfassen, Umformulieren und Instruction-Following. Unsloth-GGUF-Build, Llama-3.2-Community-Lizenz, vollständig offline betreibbar.

Meta Version 3.2 Kommerzielle Nutzung erlaubt Dense 3.21 B (3.21 B aktiv) 128 K Context 12/2023 local getestet

  • Restricted Weights
  • Nano
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Restricted-Weights
  • Real-Time

Sovereign Risk: LOW TODO

Schlüsselmetriken

Score · Latenz · Kosten · Qualität

Total Score Bronze
52.05
Routine
32.2
Reasoning
19.86

Rank #98

LLM Judge Avg
2.37
100 Coverage
Avg Task Duration
8.46
Real-Time
Token Rate
55.02
Output Rate
P95 Latency
17.18
Top 5 %
Total Tokens
52200
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 52200 tok

Benchmark-Module

10 Module · gewichtet · vs. Modellmedian & Spitzenreiter

Llama 3.2 3B (Unsloth) Bestes Modell Ø Alle Modelle
Code Quality 46.8
CLI Benchmark 70.01
Logical Reasoning 45.05
UX Writing 54.15
Documentation 43.29
Content Transform. 60.97
Cultural Intelligence 57.3
Synthesis Quality 31.67
Tool Execution 63.33
ToolUse Score 47.83
Benchmark Cost $0

Token-Effizienz & Latenz

Verbrauch pro Modul vs. Modellmedian

Token-Verbrauch pro Modul

Performance-Profil