Can Qwen3.8-Flash-Next run on MacBook Air M5 24GB?
NO — Won’t fit
IQ4_XS · 4K context · formula estimate
What this means
The checked IQ4_XS build needs about 107.9GB, above this Mac's conservative 21GB working budget.
−86.9GB short
above the working budget
107.9GB
needed at IQ4_XS
IQ4_XS
quant selected
0GB
working headroom
needs 107.9 GBworking budget 21 GB · short 86.9 GB
Where the memory goes
| Component | Detail | GB |
|---|---|---|
| Model weights | IQ4_XS GGUF (93.7 GB) + mmap overhead | 98.4 |
| KV cache | 4K context window | 8.8 |
| Runtime | macOS inference app + compute buffers | 0.8 |
| Total needed | at IQ4_XS, 4K context | 107.9 |
| Working budget | 24 GB unified memory − conservative macOS reserve | 21 |
| Headroom | memory shortfall | −86.9 |
Pick your quant
| Quant | Download | Memory | Estimated speed | Verdict |
|---|---|---|---|---|
| IQ1_S | 72.5 GB | 85.7 GB | — | ✕ Won't fit |
| IQ4_XS ★ | 93.7 GB | 107.9 GB | — | ✕ Won't fit |
Other models on this Mac
Qwen3.8-Flash-Next on other MacBook Air M5 configurations
FAQ
Can the MacBook Air M5 · 24GB run Qwen3.8-Flash-Next?
Not with the quants currently tracked. The selected IQ4_XS build needs about 107.9GB, above the 21GB working budget.
Which Qwen3.8-Flash-Next quant should I use on this Mac?
None of the tracked quants fit this configuration safely. Choose a smaller model or a Mac with more unified memory.
Which app should I use for Qwen3.8-Flash-Next on this Mac?
Current beginner Mac apps do not load this exact file. Use Bonsai 27B 1-bit in Locally AI instead, or open the advanced publisher instructions on this page.
Are these speeds measured on a MacBook Air M5?
No. The range is a formula estimate based on memory bandwidth, model size, active parameters, cooling, and a 4K context. The app, backend, thermals, and prompt can change real performance.