video model · wan · Windows
Can I run Wan 2.1 T2V 1.3B on Nvidia GeForce RTX 2060 (6GB)?
Needs ~6 GB at Q4 GGUF, but only ~5 GB is usable on Nvidia GeForce RTX 2060 (6GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
Needs ~6 GB at Q4 GGUF, but only ~5 GB is usable on Nvidia GeForce RTX 2060 (6GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
The gap is about 1 GB: Wan 2.1 T2V 1.3B needs roughly 6 GB and Nvidia GeForce RTX 2060 (6GB) leaves only about 5 GB usable for a model. The lightest tracked hardware that runs Wan 2.1 T2V 1.3B is the Nvidia GeForce RTX 3060 Ti (8GB) at 8 GB. See Wan 2.1 T2V 1.3B on Nvidia GeForce RTX 3060 Ti (8GB).
- Peak VRAM
- ~6 GB
- Usable on device
- ~5 GB
- Device memory
- 6 GB
- Quant
- Q4 GGUF
- Type
- video (DIT)
- Parameters
- 1.3B
- Peak VRAM
- ~6 GB at Q4 GGUF
- Resolution
- 832×480 (480p)
- License
- Apache-2.0
- Memory
- 6 GB vram
- Usable for weights
- ~5 GB
- Power draw
- ~160 W
- Best runtime
- Ollama (CUDA) / llama.cpp CUDA
Run Wan 2.1 T2V 1.3B on other hardware
FAQ
Can Nvidia GeForce RTX 2060 (6GB) run Wan 2.1 T2V 1.3B?
Needs ~6 GB at Q4 GGUF, but only ~5 GB is usable on Nvidia GeForce RTX 2060 (6GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
How much VRAM does Wan 2.1 T2V 1.3B need?
Nvidia GeForce RTX 2060 (6GB) does not have enough memory. At Q4 GGUF the realistic peak is ~6 GB of VRAM, versus ~20 GB with every component kept resident (no offload). With aggressive CPU offload it drops to ~5 GB, much slower.
What do I use to run Wan 2.1 T2V 1.3B locally?
Wan 2.1 T2V 1.3B runs in ComfyUI or Diffusers. It loads as a video diffusion checkpoint plus its text encoder and VAE, not a single chat command.
Sources
VRAM figures are sourced peak-usage anchors at the noted quant, validated 2026-09-14. See methodology.