Qwen 3.6 35B-A3B (Uncensored)
This community fine-tune variant of Qwen 3.6 35B-A3B removes the safety filters and delivers unfiltered responses without Refusals. Of the 35 billion total parameters in the MoE architecture, only 3 billion are active per token; the context window spans 262,000 tokens. Operable locally at near-full quality with Q8 quantization under the Apache 2.0 license, with multimodal processing for text, image, and video.
- Open Weights
- Workstation
- llama.cpp
- Text
- Vision
- Video
- Instruction-Tuned
- Uncensored
- Agentic Orchestrator
- Interactive
Sovereign Risk: HIGH Community fine-tune of a Chinese Open Weights model. The weights are publicly available on Hugging Face. Due to the Chinese origin and the National Security Law (NSL), there is an elevated risk. Additionally, this is an uncensored variant without safety filters.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 71.69
- Routine
- 44.52
- Reasoning
- 27.17
- LLM Judge Avg
- 3.62 / 5
- 100 Coverage
- Avg Task Duration
- 24.18s
- Interactive
- Token Rate
- 53.85tok/s
- Output Rate
- P95 Latency
- 78.05s
- Top 5 %
- Total Tokens
- 78200
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 78200 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median