Gemma 3 270M (Unsloth)

With 270 million parameters, Gemma 3 270M is the smallest Gemma 3 model and a text-only model for local latency baselines and embedded setups. The Unsloth GGUF variant runs fully offline under the Gemma terms of use, but is not suitable as a general-purpose quality anchor.

Google Version 3 Commercial use permitted Dense 0.27 B (0.27 B active) 32 K Context 12/2024 locally tested

  • Restricted Weights
  • Nano
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Restricted-Weights
  • Real-Time

Sovereign Risk: LOW TODO

Key metrics

Score · Latency · Cost · Quality

Total Score Standard
23.59
Routine
15.27
Reasoning
8.31

Rank #101

LLM Judge Avg
0.86
100 Coverage
Avg Task Duration
4.18
Real-Time
Token Rate
211.2
Output Rate
P95 Latency
5.65
Top 5 %
Total Tokens
75400
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 75400 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Gemma 3 270M (Unsloth) Best model Ø All models
Code Quality 15.2
CLI Benchmark 51.12
Logical Reasoning 17.45
UX Writing 19.85
Documentation 10.04
Content Transform. 33.8
Cultural Intelligence 31.4
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile