AI Models for MacBook Air M5 — What runs on 32GB

60 great · 1 slow · 5 won't fit
Chip
Apple M5
Memory bandwidth
153 GB/s
Unified memory
32 GB
Usable for models
~28.8 GB
Cooling
Fanless
Year
2026

Specs checked against manufacturer and public documentation on . Results below are estimates, not measurements.

What runs on the MacBook Air M5

All 66 models at the best tracked quant for this 32GB configuration. Select any row for the full report and step-by-step setup guide.

ModelParamsQuantNeedsSpeedVerdict

~ = bandwidth-based estimate · open a row to see the recommended app and exact steps

Other MacBook Air M5 memory options

FAQ

What is the biggest local AI model the MacBook Air M5 · 32GB can run?

Qwen 3.6 35B-A3B is the largest model in our 66-model comparison that fits at Q4_K_M. It needs about 24.1GB inside our 28.8GB working-memory budget, with an estimated 33.9–46.8 tokens/s decode range.

How much of the 32GB unified memory is available to a local LLM?

We use a conservative 28.8GB working budget, leaving room for macOS, the inference app, and normal background activity. Memory pressure, context length, and other open apps can change the real limit.

Can the MacBook Air M5 · 32GB run Llama 3.1 8B?

Yes. At Q4_K_M, our estimate uses 6.5GB and lands around 13.1–18.1 tokens/s.

Can the MacBook Air M5 · 32GB run Qwen 3.6 27B?

Yes at Q4_K_M: about 18.7GB of working memory and an estimated 3.8–5.3 tokens/s.

Which app should I use for local AI on the MacBook Air M5?

For an exact GGUF repository and quant from this site, start with LM Studio's graphical Discover, download, load, and chat flow. Jan is the open-source GGUF alternative; Ollama is strongest when the exact model already has a trustworthy catalog package or you want coding/API integrations; Msty is useful for mixed GGUF, MLX, and document workflows. For the curated Bonsai 27B build, try Locally AI and first confirm that its catalog shows the model on this Mac.

Can I upgrade the MacBook Air M5 to more unified memory later?

No. Apple-silicon unified memory is integrated into the chip package and must be chosen at purchase. Storage upgrades or external SSDs do not increase the memory an LLM can use.

Are these MacBook Air M5 local AI speeds measured?

No. Every speed on this page is a formula range based on Apple’s published memory bandwidth, model size, active parameters, and a cooling-aware efficiency range. App, backend, context length, and thermals can change real results.