Models /gpt-oss-120b /Can I run it?
Compatibility check

Can I run gpt-oss-120b on Apple M2 Max (38-core GPU)?

Probably partially, with CPU offload.best quant: Q4_K_M

This one is close. The answer rests on a file size we calculated from the parameter count rather than measured. A 10% error in that size changes the verdict, so treat it as a maybe and test before you buy.

Decode speed
Usable context
Memory
96 GB
Q4_K_MSpills to system RAMest
Q5_K_MSpills to system RAM
Q8_0Too large
Plan B — rent it

gpt-oss-120b is hosted by 18 providers from $0.17 per 1M output tokens (OpenRouter).Compare providers →

gpt-oss-120b
120B · full specs & benchmarks →
Apple M2 Max (38-core GPU)
96 GB · everything it runs →