Qwen 3.6 35B-A3B (Uncensored)

This community fine-tune variant of Qwen 3.6 35B-A3B removes the safety filters and delivers unfiltered responses without Refusals. Of the 35 billion total parameters in the MoE architecture, only 3 billion are active per token; the context window spans 262,000 tokens. Operable locally at near-full quality with Q8 quantization under the Apache 2.0 license, with multimodal processing for text, image, and video.

Alibaba Version 3.6 Commercial use permitted MoE 35 B (3 B active) 262 K Context 06/2025 locally tested

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Vision
  • Video
  • Instruction-Tuned
  • Uncensored
  • Agentic Orchestrator
  • Interactive

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
71.69
Routine
44.52
Reasoning
27.17

Rank #56

LLM Judge Avg
3.62
100 Coverage
Avg Task Duration
24.18
Interactive
Token Rate
53.85
Output Rate
P95 Latency
78.05
Top 5 %
Total Tokens
78200
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 78200 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Qwen 3.6 35B-A3B (Uncensored) Best model Ø All models
Code Quality 74.3
CLI Benchmark 90.56
Logical Reasoning 67.07
UX Writing 72.05
Documentation 73.31
Content Transform. 70.48
Cultural Intelligence 75.6
Synthesis Quality 40.83
Tool Execution 81.67
ToolUse Score 59.58
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile