Compatibility check
Can I run Qwen3.5-122B-A10B on Apple M2 Max (38-core GPU)?
Partially — with CPU offload.best quant:
78.9 GB of weights, plus 3.8 GB for the software that runs it and the smallest conversation it can hold, comes to 82.7 GB against the 72 GB this 96 GB device leaves free.
Decode speed
—
Usable context
—
Memory
96 GB
Spills to system RAM
Spills to system RAM
Too large
Plan B — rent it
Qwen3.5-122B-A10B is hosted by 6 providers from $2.08 per 1M output tokens (Alibaba Cloud).Compare providers →