Nemotron 3 Nano 30B-A3B

At the default Q4_K_M, 0 of 140 phones have enough estimated usable memory; 0 also reach 3+ tokens/s. On Mac, 36 of 54 exact configurations pass using that quant or a smaller tracked fallback.

Params
30B (3B active)
Family
nemotron
Released
2025-12
Tags
chat · reasoning

Downloads & memory by quant

The download is only the weights — add KV cache and runtime to get what your phone actually needs. “Memory-fit” only means it can load; “usable” also requires an estimated 3+ tokens/s. Each row uses that row's quant. ★ = recommended.

QuantDownloadMemory neededPhone reality
IQ4_XS18.2 GB~21.8 GB0 memory-fit
Q4_018.2 GB~21.8 GB0 memory-fit
Q3_K_M20 GB~23.7 GB0 memory-fit
Q4_K_S22 GB~25.8 GB0 memory-fit
Q4_K_M24.6 GB~28.5 GB0 memory-fit
Q5_K_M26.1 GB~30.1 GB0 memory-fit
Q6_K33.5 GB~37.9 GB0 memory-fit
Q8_033.6 GB~38 GB0 memory-fit

Nemotron 3 Nano 30B-A3B at Q4_K_M on 140 phones

This table is fixed to the recommended Q4_K_M, on each phone's largest RAM option. Counts for smaller quants above will differ. All phones are shown, sorted by fit and estimated experience; tap one for the full report.

PhoneChip · RAMNeedsSpeedVerdict

~ = bandwidth-based estimate · ✓ = measured on real hardware

Nemotron 3 Nano 30B-A3B on 54 Mac configurations

Each Mac uses the recommended Q4_K_M at 4K context when it fits, or the largest smaller tracked quant that fits. Each row links to the full report with a macOS-specific setup guide.

MacUnified memoryQuantWorking budgetNeedsEstimated speedVerdict

FAQ

How much RAM do I need to run Nemotron 3 Nano 30B-A3B on a phone?

At Q4_K_M, Nemotron 3 Nano 30B-A3B needs ~28.5GB of usable memory (weights + KV cache + runtime). In practice that means no current Android phone or no current iPhone — iOS lets a single app use less of its RAM than Android does.

How fast does Nemotron 3 Nano 30B-A3B run on a flagship phone?

It doesn't fit on any of the 140 phones we track at Q4_K_M — it needs ~28.5GB of usable memory.

Can an iPhone run Nemotron 3 Nano 30B-A3B?

Not really: the best iPhone we track (iPhone 16 Pro Max) has ~5.2GB usable, but Nemotron 3 Nano 30B-A3B needs 28.5GB at Q4_K_M.

What is the best quantization of Nemotron 3 Nano 30B-A3B for mobile?

Q4_K_M (24.6GB download) is the size/quality sweet spot of the 8 quants available. Total memory needed is ~28.5GB once the KV cache and runtime are counted — the download size alone understates it.

How many phones can run Nemotron 3 Nano 30B-A3B?

0 of the 140 phones we track have enough estimated usable memory at Q4_K_M; 0 also reach our usable-speed threshold of 3 tokens/s, and 0 run smoothly. Memory-fit alone does not mean a good experience.

Which Macs can run Nemotron 3 Nano 30B-A3B?

36 of the 54 exact Mac configurations we track have enough estimated working memory. Each Mac uses the recommended Q4_K_M when it fits, or the largest smaller tracked quant that fits, and separates installed unified memory, the macOS reserve, and estimated speed.

What is Nemotron 3 Nano 30B-A3B good for on a phone?

It's tagged for chat, reasoning. At 30B parameters (3B active — it's a MoE, so it decodes faster than its size suggests), it prioritizes answer quality over speed — expect slower decoding.

Similar models

Nemotron 3 Nano 4BNemotron 3 Nano Omni 30B-A3BNemotron 3.5 Lightning 30B-A3BSalience 1.5 FlashGranite 4.2 30B