Can OneJev 27B run on ROG Phone 9 Pro?

Estimated weights fit · custom workflow required

Weight fit is not phone app support

The tracked Q4_K_M weights are 16.5GB; the generic working-memory estimate is 19.8GB against 20GB usable. Image inputs need additional projector memory.

OneJev scores typed decisions in a single forward pass, not chat replies. The decode estimates below are not measured decision latency. The publisher's qev server and System One client are required; AICanRun has not verified a phone app route.

Not measured
decision latency
19.8GB
needed at Q4_K_M
Q4_K_M
quant checked
needs 19.8 GBusable 20 GB
Snapdragon 8 Elite24 GB RAMCPU / GPUAlso sold with 16 GB

Make it bearable

1
Drop context to 2K — saves ~1 GB of KV cache and a little speed.
2
Close every other app before loading — the 0.2 GB headroom is real; Android may kill the app otherwise.
3
Expect throttling after ~10 min of sustained generation on a phone chassis.

Where the memory goes

ComponentDetailGB
Model weightsQ4_K_M GGUF (16.5 GB) + mmap overhead17.3
KV cache4K context window1.9
Runtimellama.cpp + app overhead0.6
Total neededat Q4_K_M, 4K context19.8
Working budget24 GB RAM − Android system reserve20
Headroomremaining inside the working budget0.2

Pick your quant

QuantDownloadVerdictSpeed
Q2_K 10.7 GB! Runs, barely~3.2 tokens/s
Q3_K_M 13.3 GB! Runs, barely~2.6 tokens/s
IQ4_XS 15.2 GB! Runs, barely~2.3 tokens/s
Q4_K_S 15.6 GB! Runs, barely~2.2 tokens/s
Q4_K_M BEST HERE16.5 GB! Runs, barely~2.1 tokens/s
Q5_K_M 19.2 GB✕ Won't fitwon't fit
Q6_K 22.1 GB✕ Won't fitwon't fit
Q8_0 28.6 GB✕ Won't fitwon't fit

Use the right workflow for this model

The model fits in memory, but it is not an ordinary chat model.

This model is designed to return calibrated probabilities for typed choices from text or images. PocketPal does not provide the required the publisher's qev server and System One client; image inputs also need a vision projector, so the tracked file is not a beginner phone-chat path.

Read the publisher workflow ↗

If you only want a normal private chat, use this simpler model instead:

Check Llama 3.2 3B

Related checks

More on ROG Phone 9 Pro
Qwen3 0.6BQwen3 1.7BLlama 3.2 1BLlama 3.2 3BGemma 3 1B
OneJev 27B on other phones
OnePlus 13OnePlus 12Huawei Mate X7Galaxy S25 UltraPixel 9 Pro