Gemma 3 270M (Unsloth)
With 270 million parameters, Gemma 3 270M is the smallest Gemma 3 model and a text-only model for local latency baselines and embedded setups. The Unsloth GGUF variant runs fully offline under the Gemma terms of use, but is not suitable as a general-purpose quality anchor.
- Restricted Weights
- Nano
- llama.cpp
- Text
- Instruction-Tuned
- Restricted-Weights
- Real-Time
Sovereign Risk: LOW TODO
Key metrics
Score · Latency · Cost · Quality
- Total Score Standard
- 23.59
- Routine
- 15.27
- Reasoning
- 8.31
- LLM Judge Avg
- 0.86 / 5
- 100 Coverage
- Avg Task Duration
- 4.18s
- Real-Time
- Token Rate
- 211.2tok/s
- Output Rate
- P95 Latency
- 5.65s
- Top 5 %
- Total Tokens
- 75400
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 75400 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Gemma 3 270M (Unsloth)
Best model
Ø All models
Code Quality
15.2
CLI Benchmark
51.12
Logical Reasoning
17.45
UX Writing
19.85
Documentation
10.04
Content Transform.
33.8
Cultural Intelligence
31.4
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost
$0
Token efficiency & latency
Consumption per module vs. model median