Models /Mixtral 8x22B Instruct /Can I run it?
Compatibility check

Can I run Mixtral 8x22B Instruct on Apple M2 Max (38-core GPU)?

Partially — with CPU offload.best quant:

88.7 GB of weights, plus 4.3 GB for the software that runs it and the smallest conversation it can hold, comes to 93 GB against the 72 GB this 96 GB device leaves free.

Decode speed
—
Usable context
—
Memory
96 GB
Spills to system RAM
Too largeest
Too large
What is quantisation? →
Plan B — rent it

Mixtral 8x22B Instruct is hosted by 2 providers from $6.00 per 1M output tokens (Mistral AI).Compare providers →

Mixtral 8x22B Instruct
141B / 39B · full specs & benchmarks →
Apple M2 Max (38-core GPU)
96 GB · everything it runs →