Qwen 3.6 35B-A3B (Uncensored)
This community fine-tune variant of Qwen 3.6 35B-A3B removes the safety filters and delivers unfiltered responses without Refusals. Of the 35 billion total parameters in the MoE architecture, only 3 billion are active per token; the context window spans 262,000 tokens. Operable locally at near-full quality with Q8 quantization under the Apache 2.0 license, with multimodal processing for text, image, and video.
- Open Weights
- Desktop
- llama.cpp
- Text
- Vision
- Video
- Instruction-Tuned
- Uncensored
- Agentic Orchestrator
- Interactive
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 71.69
- Routine
- 44.52
- Reasoning
- 27.17
- LLM Judge Avg
- 3.62 / 5
- 100 Coverage
- Avg Task Duration
- 24.18s
- Interactive
- Token Rate
- 53.85tok/s
- Output Rate
- P95 Latency
- 78.05s
- Top 5 %
- Total Tokens
- 78200
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 78200 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median