Compatibility check
Can I run Qwen3.5-35B-A3B on Apple M3 Pro (18-core GPU)?
Yes, but only just.best quant: Q4_K_M
This one is close. The answer rests on a file size we calculated from the parameter count rather than measured, and it leaves 2.3 GB spare once the weights, a 2K context and runtime overhead are counted. A 10% error in that size changes the verdict, so treat it as a maybe and test before you buy.
Decode speed
49 tok/sest
Usable context
16K
Memory
36 GB
Q4_K_MFits in memoryest49 tok/sest · 16K ctx
Q5_K_MSpills to system RAMest
Q8_0Too largeest