Compatibility check
Can I run Gemma 4 E4B it on Apple M2 (8-core GPU, 8GB unified)?
Partially — with CPU offload.best quant:
5 GB of weights, plus 1.5 GB for the software that runs it and the smallest conversation it can hold, comes to 6.5 GB against the 6 GB this 8 GB device leaves free.
Decode speed
—
Usable context
—
Memory
8 GB
Spills to system RAM
Spills to system RAM
Too large
Plan B — rent it
Gemma 4 E4B it is hosted by 1 provider from $0.10 per 1M output tokens (DeepInfra).Compare providers →