Qwen3 0.6B

At the default Q8_0, 140 of 140 phones have enough estimated usable memory; 140 also reach 3+ tokens/s. On Mac, 54 of 54 exact configurations pass using that quant or a smaller tracked fallback.

Params
0.6B
Family
qwen
Released
2025-04
Tags
chat

Downloads & memory by quant

The download is only the weights — add KV cache and runtime to get what your phone actually needs. “Memory-fit” only means it can load; “usable” also requires an estimated 3+ tokens/s. Each row uses that row's quant. ★ = recommended.

QuantDownloadMemory neededPhone reality
Q8_00.6 GB~1.3 GB140 memory-fit · 140 usable

Qwen3 0.6B at Q8_0 on 140 phones

This table is fixed to the recommended Q8_0, on each phone's largest RAM option. Counts for smaller quants above will differ. All phones are shown, sorted by fit and estimated experience; tap one for the full report.

PhoneChip · RAMNeedsSpeedVerdict

~ = bandwidth-based estimate · ✓ = measured on real hardware

Qwen3 0.6B on 54 Mac configurations

Each Mac uses the recommended Q8_0 at 4K context when it fits, or the largest smaller tracked quant that fits. Each row links to the full report with a macOS-specific setup guide.

MacUnified memoryQuantWorking budgetNeedsEstimated speedVerdict

FAQ

How much RAM do I need to run Qwen3 0.6B on a phone?

At Q8_0, Qwen3 0.6B needs ~1.3GB of usable memory (weights + KV cache + runtime). In practice that means a 6GB+ Android phone or a 6GB+ iPhone — iOS lets a single app use less of its RAM than Android does.

How fast does Qwen3 0.6B run on a flagship phone?

On the Galaxy S25 Ultra (Snapdragon 8 Elite) it runs at ~57.6 tokens/s at Q8_0, estimated from memory bandwidth. Anything above ~8 tokens/s feels smooth for chat.

Can an iPhone run Qwen3 0.6B?

Yes — the iPhone 16 Pro Max runs it at ~45 tokens/s at Q8_0, using 1.3GB of its ~5.2GB usable memory.

What is the best quantization of Qwen3 0.6B for mobile?

Q8_0 (0.6GB download) is the size/quality sweet spot. Total memory needed is ~1.3GB once the KV cache and runtime are counted — the download size alone understates it.

How many phones can run Qwen3 0.6B?

140 of the 140 phones we track have enough estimated usable memory at Q8_0; 140 also reach our usable-speed threshold of 3 tokens/s, and 140 run smoothly. Memory-fit alone does not mean a good experience.

Which Macs can run Qwen3 0.6B?

54 of the 54 exact Mac configurations we track have enough estimated working memory. Each Mac uses the recommended Q8_0 when it fits, or the largest smaller tracked quant that fits, and separates installed unified memory, the macOS reserve, and estimated speed.

What is Qwen3 0.6B good for on a phone?

It's tagged for chat. At 0.6B parameters, it's a fast, lightweight pick for quick tasks.

Guides for Qwen3

Bonsai 27B vs Qwen3.6 27B on Phones: Is 1-Bit Worth It?

Similar models

Qwen3 1.7BQwen3 4BQwen3 8BQwen3 14BQwen3 30B A3B