LensVLM 9B

At the default Q4_K_M, 111 of 148 phones have enough estimated usable memory; 95 also reach 3+ tokens/s. On Mac, 52 of 54 exact configurations pass using that quant or a smaller tracked fallback.

Params
9.4B
Family
lensvlm
Released
2026-09
Tags
chat · vision

Downloads & memory by quant

The download is only the weights — add KV cache and runtime to get what your phone actually needs. “Memory-fit” only means it can load; “usable” also requires an estimated 3+ tokens/s. Each row uses that row's quant. ★ = recommended.

QuantDownloadMemory neededPhone reality
Q2_K3.6 GB~5 GB143 memory-fit · 131 usable
Q3_K_M4.5 GB~6 GB136 memory-fit · 103 usable
IQ4_XS5.2 GB~6.7 GB111 memory-fit · 96 usable
Q4_05.5 GB~7 GB111 memory-fit · 95 usable
Q4_K_S5.5 GB~7 GB111 memory-fit · 95 usable
Q4_K_M ★5.8 GB~7.3 GB111 memory-fit · 95 usable
Q5_K_M6.9 GB~8.5 GB108 memory-fit · 92 usable
Q6_K7.8 GB~9.4 GB64 memory-fit · 64 usable
Q8_09.5 GB~11.2 GB64 memory-fit · 64 usable

LensVLM 9B at Q4_K_M on 148 phones

This table is fixed to the recommended Q4_K_M, on each phone's largest RAM option. Counts for smaller quants above will differ. All phones are shown, sorted by fit and estimated experience; tap one for the full report.

PhoneChip · RAMNeedsSpeedVerdict

~ = bandwidth-based estimate · ✓ = measured on real hardware

LensVLM 9B on 54 Mac configurations

Each Mac uses the recommended Q4_K_M at 4K context when it fits, or the largest smaller tracked quant that fits. Each row links to the full report with a macOS-specific setup guide.

MacUnified memoryQuantWorking budgetNeedsEstimated speedVerdict

FAQ

How much RAM do I need to run LensVLM 9B on a phone?

At Q4_K_M, LensVLM 9B needs ~7.3GB of usable memory (weights + KV cache + runtime). In practice that means a 12GB+ Android phone or a 12GB+ iPhone — iOS lets a single app use less of its RAM than Android does.

How fast does LensVLM 9B run on a flagship phone?

On the Xiaomi 18 Fold (Xiaomi XRING O3) it runs at ~8.8 tokens/s at Q4_K_M, estimated from memory bandwidth. Anything above ~8 tokens/s feels smooth for chat.

Can an iPhone run LensVLM 9B?

Yes — the iPhone Air runs it at ~6 tokens/s at Q4_K_M, using 7.3GB of its ~7.8GB usable memory.

What is the best quantization of LensVLM 9B for mobile?

Q4_K_M (5.8GB download) is the size/quality sweet spot of the 9 quants available. Total memory needed is ~7.3GB once the KV cache and runtime are counted — the download size alone understates it.

How many phones can run LensVLM 9B?

111 of the 148 phones we track have enough estimated usable memory at Q4_K_M; 95 also reach our usable-speed threshold of 3 tokens/s, and 1 run smoothly. Memory-fit alone does not mean a good experience.

Which Macs can run LensVLM 9B?

52 of the 54 exact Mac configurations we track have enough estimated working memory. Each Mac uses the recommended Q4_K_M when it fits, or the largest smaller tracked quant that fits, and separates installed unified memory, the macOS reserve, and estimated speed.

What is LensVLM 9B good for on a phone?

It's tagged for chat, vision. At 9.4B parameters, it prioritizes answer quality over speed — expect slower decoding.

Similar models

MiMo V2.6 Distill Qwen 9BGRM 3.2 Cliff 9Bgrug 9B (ProCreations)Fara 1.5 9BQwen 3.5 9B