Compatibility check
Can I run Mistral Small 4 on Apple M2 Max (38-core GPU)?
Partially — with CPU offload.best quant:
75.3 GB of weights, plus 4.7 GB for the software that runs it and the smallest conversation it can hold, comes to 80 GB against the 72 GB this 96 GB device leaves free.
Decode speed
—
Usable context
—
Memory
96 GB
Spills to system RAM
Spills to system RAM
Too large
Plan B — rent it
Mistral Small 4 is hosted by 2 providers from $0.60 per 1M output tokens (Mistral AI).Compare providers →