Compatibility check
Can I run Mixtral 8x22B Instruct on Apple M4 Max (40-core GPU)?
Yes, but only just.best quant:
This one is close. The answer rests on a file size we calculated from the parameter count rather than measured, and it leaves 3 GB spare once the weights, a 2K context and runtime overhead are counted. A 10% error in that size changes the verdict, so treat it as a maybe and test before you buy.
Decode speed
14 tok/sest
Usable context
8K
Memory
128 GB
Fits in memoryest14 tok/sest · 8K ctx
Spills to system RAM
Too large