Skip to content
localmodel.run

Model family · 4 sizes

DeepSeek-R1-Distill: which size runs locally?

DeepSeek-R1-Distill comes in 4 sizes, from 7B to 32B. Bigger is generally more capable but needs more memory. Here is each size with its Q4_K_M weight, the memory it needs, and the hardware that runs it.

Sizes
4
Smallest
7B
Largest
32B
Runs from
16GB

The DeepSeek-R1-Distill lineup

"Needs" is the sourced minimum memory for Q4_K_M with a small context. Larger context needs more.

Which DeepSeek-R1-Distill fits your memory

8GB

No DeepSeek-R1-Distill size fits 8GB; even DeepSeek-R1-Distill-Qwen 7B needs more.

No
16GB

Largest that fits: DeepSeek-R1-Distill-Qwen 14B (14B), best case on Nvidia GeForce RTX 4080 (16GB).

Yes
24GB

Largest that fits: DeepSeek-R1-Distill-Qwen 32B (32B), best case on Nvidia GeForce RTX 4090 (24GB). Comfortable up to DeepSeek-R1-Distill-Qwen 14B (14B).

Tight
32GB

Largest that fits: DeepSeek-R1-Distill-Qwen 32B (32B), best case on Nvidia GeForce RTX 5090 (32GB).

Yes

Best case means the most capable device at that size (usually a discrete GPU). A Mac at the same size sits roughly one rung lower; see the per-size breakdown on each memory budget page.

FAQ

Which DeepSeek-R1-Distill size should I run locally?

Pick the largest size your memory allows. On 16GB (best case) up to DeepSeek-R1-Distill-Qwen 14B; On 24GB (best case) up to DeepSeek-R1-Distill-Qwen 32B; On 32GB (best case) up to DeepSeek-R1-Distill-Qwen 32B. Smaller sizes run faster and leave headroom for context.

What is the smallest DeepSeek-R1-Distill model?

DeepSeek-R1-Distill-Qwen 7B at 7B parameters, about 4.68 GB on disk at Q4_K_M and roughly 6 GB of memory to run. It is the one to use on phones and 8 GB machines.

What is the largest DeepSeek-R1-Distill model and what does it need?

DeepSeek-R1-Distill-Qwen 32B at 32B, about 19.85 GB at Q4_K_M and roughly 22 GB of memory. It fits a high-memory desktop GPU or Mac.

Sources

Memory figures are estimates at Q4_K_M. See methodology.