Qwen 3.6 35B-A3B

Apache 2.0

Alibaba Β· 36B (3B active) Β· Mixture of Experts

Big-model quality at 3B-active speed β€” the mid-hardware sweet spot Check if your GPU or Mac can run Qwen 3.6 35B-A3B locally β€” 20.1 GB min, 33.5 GB recommended.

2026-04256K context

Mixture of Experts

Total experts: 256
Active experts: 8
Active params: 3.0B

Quantization Options

QuantBitsVRAMQualityStatus
Q2_K212 GBlowβ€”
Q3_K_M316.6 GBmoderateβ€”
Q4_K_M418.9 GBgoodβ€”
Q5_K_M523.6 GBgoodβ€”
Q6_K628.2 GBexcellentβ€”
Q8_0837.4 GBexcellentβ€”
F161674.3 GBlosslessβ€”

Can I run Qwen 3.6 35B-A3B locally?

Can I run Qwen 3.6 35B-A3B locally?
Qwen 3.6 35B-A3B needs about 20.1 GB of memory at a minimum and 33.5 GB recommended. Open this page to grade it against your GPU or Mac, then run it with runai, Ollama or LM Studio.
How much VRAM does Qwen 3.6 35B-A3B need?
At Q4_K_M, Qwen 3.6 35B-A3B uses about 18.9 GB of VRAM. Higher quants need more memory; lower quants fit tighter cards with a quality tradeoff.