Signal 3.8 27B
Signal 3.8 27B by AgentionAI is a token-efficiency fine-tune on Qwen 3.8 27B: via self-distillation, the model generates roughly 57 percent fewer response tokens and thinking chains half the length, at equal or better quality. Licensed under Apache 2.0, available in a mid-range Q5 quantization with a 262,000-token context and local inference.
- Open Weights
- Workstation
- llama.cpp
- Text
- Vision
- Instruction-Tuned
- Batch
Sovereign Risk: MEDIUM TODO
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.23
- Routine
- 44.13
- Reasoning
- 29.1
- LLM Judge Avg
- 3.63 / 5
- 100 Coverage
- Avg Task Duration
- 52.67s
- Batch
- Token Rate
- 14.49tok/s
- Output Rate
- P95 Latency
- 158.13s
- Top 5 %
- Total Tokens
- 69300
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 69300 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median