Qwen
Qwen
/Qwen3.6-35B-A3B

Quantizations

QuantQuantized bySizeDecodePrefillScoreActions
Unsloth
Unsloth
9.4 GBN/AN/AN/A
Unsloth
Unsloth
10.0 GBN/AN/AN/A
Unsloth
Unsloth
10.7 GBN/AN/AN/A
Frank Denis
Frank Denis
14.1 GBN/AN/AN/A
DuoNeural
DuoNeural
17.4 GB73.5 tok/s1,504.0 tok/sRuns well
MLX Community
MLX Community
19.0 GB95.8 tok/s2,360.1 tok/sRuns well
MLX Community
MLX Community
19.0 GB90.2 tok/s2,150.3 tok/sRuns well
MLX Community
MLX Community
19.2 GBN/AN/AN/A
Unsloth
Unsloth
19.5 GBN/AN/AN/A
LM Studio
LM Studio
19.7 GBN/AN/AN/A
DuoNeural
DuoNeural
19.7 GB71.5 tok/s1,532.6 tok/sRuns well
LM Studio
LM Studio
27.1 GBN/AN/AN/A
MLX Community
MLX Community
35.1 GBN/AN/AN/A
Unsloth
Unsloth
35.8 GBN/AN/AN/A

Device Comparison

Results include trials with 4,096 input tokens and 1,024 output tokens only.

Decode / Prefill Speeds

25 devices