Can You Run Gemma 4 on a MacBook Air (8GB)?
Yes — best fit is Gemma 4 E2B at Q4_0 (4-bit) (~2.5GB) ⚠️ Tight, using 3GB of usable memory on a 8GB MacBook Air.
Every Gemma 4 model on this configuration
| Model | Best quant | Size | Status |
|---|---|---|---|
| Gemma 4 E2B | Q4_0 (4-bit) | ~2.5 GB | ⚠️ Tight |
| Gemma 4 E4B | — | — | ❌ Won't fit |
| Gemma 4 12B | — | — | ❌ Won't fit |
| Gemma 4 26B A4B | — | — | ❌ Won't fit |
| Gemma 4 31B | — | — | ❌ Won't fit |
How we got this verdict
8 - 5 = 3GB usable (macOS and background processes reserve roughly 5 GB of unified memory).
E2B lands on Q4_0 here, so grab the official QAT (quantization-aware training) checkpoint instead of a generic Q4_0 quant. Google shipped a full-quality QAT build for E2B at essentially the same ~4.3GB download size, with noticeably better output quality than a naive 4-bit quant. See the QAT guide for exact file names.
Going up to 16GB (16GB) unlocks 12B at Q4_0 (4-bit) — a step up from E2B here.
Try a different configuration
Best fit: Gemma 4 E2B · Q4_0 (4-bit) · ~2.5 GB ⚠️ Tight
FAQ
Can a MacBook Air (8GB) run the 31B flagship model?
No — the 31B model needs at least 17.4GB usable memory even at Q4_0, and this configuration only has 3GB usable. Gemma 4 E2B is the largest model that fits here.
What's the best Gemma 4 model for a MacBook Air (8GB)?
Gemma 4 E2B at Q4_0 (4-bit) (~2.5GB) is the best fit — it's the largest model that runs at all (tightly) within the 3GB of usable memory here.
What if I have more or less memory than 8GB?
With more (e.g. 16GB), you can run larger models or the same model more comfortably — see the MacBook Air page. This is already close to the practical floor for running Gemma 4 at all.