Can GLM-4.5V run on MacBook Pro M5 Max 36GB?
NO — Won’t fit
Q4_K_M · 4K context · formula estimate
What this means
The checked Q4_K_M build needs about 75GB, above this Mac's conservative 32.4GB working budget.
−42.6GB short
above the working budget
75GB
needed at Q4_K_M
Q4_K_M
quant selected
0GB
working headroom
needs 75 GBworking budget 32.4 GB · short 42.6 GB
Where the memory goes
| Component | Detail | GB |
|---|---|---|
| Model weights | Q4_K_M GGUF (63.6 GB) + mmap overhead | 66.8 |
| KV cache | 4K context window | 7.4 |
| Runtime | macOS inference app + compute buffers | 0.8 |
| Total needed | at Q4_K_M, 4K context | 75 |
| Working budget | 36 GB unified memory − conservative macOS reserve | 32.4 |
| Headroom | memory shortfall | −42.6 |
Pick your quant
| Quant | Download | Memory | Estimated speed | Verdict |
|---|---|---|---|---|
| Q4_K_M ★ | 63.6 GB | 75 GB | — | ✕ Won't fit |
| Q8_0 | 113.6 GB | 127.5 GB | — | ✕ Won't fit |
Other models on this Mac
GLM-4.5V on other MacBook Pro M5 Max configurations
FAQ
Can the MacBook Pro M5 Max · 36GB run GLM-4.5V?
Not with the quants currently tracked. The selected Q4_K_M build needs about 75GB, above the 32.4GB working budget.
Which GLM-4.5V quant should I use on this Mac?
None of the tracked quants fit this configuration safely. Choose a smaller model or a Mac with more unified memory.
Which app should I use for GLM-4.5V on this Mac?
Start with LM Studio. This page gives the complete point-and-click walkthrough.
Are these speeds measured on a MacBook Pro M5 Max?
No. The range is a formula estimate based on memory bandwidth, model size, active parameters, cooling, and a 4K context. The app, backend, thermals, and prompt can change real performance.