Can OneJev 9B run on iPhone 16 Pro Max?

Weights exceed this memory budget · custom workflow required

Weight fit is not phone app support

The tracked Q4_K_M weights are 5.6GB; the generic working-memory estimate is 7.1GB against 5.2GB usable. Image inputs need additional projector memory.

OneJev scores typed decisions in a single forward pass, not chat replies. The decode estimates below are not measured decision latency. The publisher's qev server and System One client are required; AICanRun has not verified a phone app route.

−1.9GB short
memory that doesn't exist here
7.1GB
needed at Q4_K_M
Q4_K_M
quant checked
needs 7.1 GBusable 5.2 GB · short 1.9 GB
Apple A18 Pro8 GB RAMMetal

Three ways forward

1
Try a smaller quant
No quant of OneJev 9B fits this phone — even the smallest is too big. Options 2 and 3 are your real choices.
2
Best model that fits
Qwen3 0.6B · Q8_0 · 0.6 GB
✓ ~45 tokens/s here
check Qwen3 0.6B →
3
Run it in the cloud
The full OneJev 9B, no download, at data-center speed — pay per token.
provider recommendations coming later

Where the memory goes

ComponentDetailGB
Model weightsQ4_K_M GGUF (5.6 GB) + mmap overhead5.9
KV cache4K context window0.6
Runtimellama.cpp + app overhead0.6
Total neededat Q4_K_M, 4K context7.1
Working budget8 GB RAM − iOS reserve (jetsam limit)5.2
Headroommemory shortfall−1.9

Pick your quant

QuantDownloadVerdictSpeed
Q2_K 3.8 GB✕ Won't fitwon't fit
Q3_K_M 4.6 GB✕ Won't fitwon't fit
IQ4_XS 5.2 GB✕ Won't fitwon't fit
Q4_K_S 5.4 GB✕ Won't fitwon't fit
Q4_K_M 5.6 GB✕ Won't fitwon't fit
Q5_K_M 6.5 GB✕ Won't fitwon't fit
Q6_K 7.4 GB✕ Won't fitwon't fit
Q8_0 9.5 GB✕ Won't fitwon't fit

Related checks

More on iPhone 16 Pro Max
Qwen3 0.6BQwen3 1.7BLlama 3.2 1BGemma 3 1BDeepSeek R1 Distill 1.5B
OneJev 9B on other phones
Xiaomi 18 FoldOPPO Find X8 Provivo X200 ProGalaxy S26 UltraXiaomi 17