AI Models for iPhone 17 Pro Max — What runs on 12GB

32 great · 15 slow · 46 won't fit
Chip
Apple A19 Pro
Memory bandwidth
76.8 GB/s
NPU
35 TOPS
RAM options
12 GB
Usable for models
~7.8 GB
Year
2025

Specs checked against manufacturer and public documentation on .

What runs on the iPhone 17 Pro Max

All 93 models at their recommended quant, on the 12GB configuration. Select any row for the full report.

ModelParamsQuantNeedsSpeedVerdict

~ = bandwidth-based estimate · ✓ = measured on real hardware

Best model by use case

Best for Chat

Top everyday assistant & writing pick here — ~57.6 tokens/s at Q8_0, using 1.3 of ~7.8GB.

Best for Coding

Top code completion & explain-this pick here — ~18.2 tokens/s at Q4_K_M, using 2.8 of ~7.8GB.

Best for Reasoning

Top math & step-by-step thinking pick here — ~31.4 tokens/s at Q4_K_M, using 1.9 of ~7.8GB.

FAQ

What is the biggest AI model the iPhone 17 Pro Max can run?

Bonsai 27B (1-bit) (27B parameters) at Q1_0 — it needs 4.8GB of the ~7.8GB usable on the 12GB iPhone 17 Pro Max and runs at ~9.1 tokens/s.

How much of the iPhone 17 Pro Max's 12GB RAM can AI models actually use?

About 7.8GB. iOS caps a single app at roughly 65% of total RAM, so of the 12GB about 7.8GB is actually available to a model.

Can the iPhone 17 Pro Max run Llama 3.1 8B?

Yes — at Q4_K_M it needs 6.3GB of the ~7.8GB usable and runs at ~7.1 tokens/s.

How fast is local AI on the iPhone 17 Pro Max?

The Apple A19 Pro has 76.8GB/s of memory bandwidth, which is what decode speed scales with. Small models like Ternary Bonsai 1.7B reach ~69.1 tokens/s; larger 7–14B models land in the single digits. Anything above ~8 tokens/s feels smooth for chat.

Which quantization should I use on the iPhone 17 Pro Max?

Q4_K_M is the size/quality sweet spot for most models. For example, Qwen3 0.6B at Q8_0 takes 1.3GB of memory here. Only drop to Q3 or IQ4 if a model just misses fitting; Q8 rarely pays off on 12GB of RAM.

Is 12GB of RAM enough for local AI?

47 of the 93 models we track fit on the iPhone 17 Pro Max — 32 run great and 15 run with compromises. 46 models (mostly 12B+) don't fit at their recommended quant.

Other Apple phones

iPhone 16 Pro MaxiPhone 16 ProiPhone 16iPhone 15 ProiPhone 15iPhone 14iPhone 17iPhone AiriPhone 17 ProiPhone 16eiPhone 17e