Compatibility check
Can I run GLM 5.2 on Apple M3 Ultra (80-core GPU)?
Partially — with CPU offload.best quant:
475 GB of weights, plus 23.3 GB for the software that runs it and the smallest conversation it can hold, comes to 498.3 GB against the 384 GB this 512 GB device leaves free.
Decode speed
—
Usable context
—
Memory
512 GB
Spills to system RAM
Too largeest
Too large
Plan B — rent it
GLM 5.2 is hosted by 27 providers from $0.91 per 1M output tokens (Reka).Compare providers →