Can XYZ-Aquila mini 35B-A3B run on MacBook Pro M3 Pro 18GB?

NO — Won’t fit
Q4_K_M · 4K context · formula estimate

What this means

The checked Q4_K_M build needs about 25.7GB, above this Mac's conservative 15GB working budget.

−10.7GB short
above the working budget
25.7GB
needed at Q4_K_M
Q4_K_M
quant selected
0GB
working headroom
needs 25.7 GBworking budget 15 GB · short 10.7 GB
Apple M3 Pro18 GB unified memoryMetalActive cooling

Where the memory goes

ComponentDetailGB
Model weightsQ4_K_M GGUF (21.4 GB) + mmap overhead22.5
KV cache4K context window2.5
RuntimemacOS inference app + compute buffers0.8
Total neededat Q4_K_M, 4K context25.7
Working budget18 GB unified memory − conservative macOS reserve15
Headroommemory shortfall−10.7

Pick your quant

QuantDownloadMemoryEstimated speedVerdict
Q2_K12.6 GB16.5 GB Won't fit
Q3_K_M16.2 GB20.3 GB Won't fit
IQ4_XS18.8 GB23 GB Won't fit
Q4_019.9 GB24.1 GB Won't fit
Q4_K_S20.6 GB24.9 GB Won't fit
Q4_K_M21.4 GB25.7 GB Won't fit
Q5_K_M25 GB29.5 GB Won't fit
Q6_K30.1 GB34.9 GB Won't fit
Q8_036.9 GB42 GB Won't fit

Other models on this Mac

XYZ-Aquila mini 35B-A3B on other MacBook Pro M3 Pro configurations

FAQ

Can the MacBook Pro M3 Pro · 18GB run XYZ-Aquila mini 35B-A3B?

Not with the quants currently tracked. The selected Q4_K_M build needs about 25.7GB, above the 15GB working budget.

Which XYZ-Aquila mini 35B-A3B quant should I use on this Mac?

None of the tracked quants fit this configuration safely. Choose a smaller model or a Mac with more unified memory.

Which app should I use for XYZ-Aquila mini 35B-A3B on this Mac?

Start with LM Studio. This page gives the complete point-and-click walkthrough.

Are these speeds measured on a MacBook Pro M3 Pro?

No. The range is a formula estimate based on memory bandwidth, model size, active parameters, cooling, and a 4K context. The app, backend, thermals, and prompt can change real performance.