Swift Qwen 3.8 27B

Swift Qwen 3.8 27B is UkisAI’s reasoning-efficiency fine-tune on Qwen 3.8 27B: up to 58 percent fewer thinking tokens at under one percent performance loss and roughly twice the throughput on reasoning tasks. NVFP4 quantization with 262,000 tokens of context, MTP head for speculative decoding, and documented tool use — license with a commercial ARR threshold.

UkisAI Version 3.8 Commercial use permitted Dense 28 B (28 B active) 262 K Context 12/2025 locally tested

  • Open Weights
  • Workstation
  • vLLM
  • Text
  • Vision
  • Gated-Weights
  • Batch

Sovereign Risk: MEDIUM TODO

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
77.78
Routine
47.67
Reasoning
30.12

Rank #11

LLM Judge Avg
4.02
100 Coverage
Avg Task Duration
106.87
Batch
Token Rate
15.52
Output Rate
P95 Latency
291.59
Top 5 %
Total Tokens
111100
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 111100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Swift Qwen 3.8 27B Best model Ø All models
Code Quality 80.6
CLI Benchmark 95
Logical Reasoning 69.87
UX Writing 73.65
Documentation 79.16
Content Transform. 74.3
Cultural Intelligence 84.3
Synthesis Quality 63.33
Tool Execution 86.67
ToolUse Score 74
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile