Models /Llama 3.3 Euryale 70B /Can I run it?
Compatibility check

Can I run Llama 3.3 Euryale 70B on Apple M4 Max (32-core GPU)?

Yes, but only just.best quant: Q4_K_M

This one is close. The answer rests on a file size we calculated from the parameter count rather than measured, and it leaves 0.3 GB spare once the weights, a 2K context and runtime overhead are counted. A 10% error in that size changes the verdict, so treat it as a maybe and test before you buy.

Decode speed
7 tok/sest
Usable context
2K
Memory
64 GB
Q4_K_MFits in memoryest7 tok/sest · 2K ctx
Q5_K_MSpills to system RAM
Q8_0Too large
Llama 3.3 Euryale 70B
70.6B · full specs & benchmarks →
Apple M4 Max (32-core GPU)
64 GB · everything it runs →