Occamy 1.0 35B-A3B (Accio-Lab) (Thinking)

Occamy 1.0 by Accio-Lab is an agentic derivative of the Qwen3.6-35B-A3B checkpoint, focused on long-horizon co-work sessions with tools, structured APIs, and persistent state tracking. The NVFP4 quantization is selective: only the routed experts are quantized, while attention, router, embeddings, and output head remain in BF16. The 35-billion-parameter MoE activates only 3 billion parameters per token and supports 262,000 tokens of context. Apache 2.0 license and documented provenance with recipe, data, and validation artifacts.

Accio-Lab Version 1.0 Commercial use permitted MoE 35 B (3 B active) 262 K Context 12/2025 locally tested

  • Open Weights
  • Workstation
  • vLLM
  • Text
  • Vision
  • Unusable

Sovereign Risk: MEDIUM Occamy 1.0 NVFP4 is a community derivative of Qwen/Qwen3.6-35B-A3B, with a published provenance trail including the quantization recipe, data-provenance file, and validation artifacts. The upstream base is Apache 2.0, the checkpoint runs locally, and the NVFP4 export is limited to routed experts, but Accio-Lab’s organizational jurisdiction is not publicly documented, so the provenance risk remains medium.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
75.25
Routine
45.68
Reasoning
29.56

Rank #32

LLM Judge Avg
3.81
100 Coverage
Avg Task Duration
124.3
Unusable
Token Rate
33.94
Output Rate
P95 Latency
379.55
Top 5 %
Total Tokens
214000
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 214000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Occamy 1.0 35B-A3B (Accio-Lab) (Thinking) Best model Ø All models
Code Quality 74.12
CLI Benchmark 83.3
Logical Reasoning 74.5
UX Writing 77.01
Documentation 74.22
Content Transform. 78.2
Cultural Intelligence 74.64
Synthesis Quality 42.5
Tool Execution 86.67
ToolUse Score 70
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile