video model · wan · Windows
Can I run Wan 2.2 TI2V 5B on Nvidia GeForce RTX 2060 (6GB)?
Needs ~8 GB at Q4 GGUF, but only ~5 GB is usable on Nvidia GeForce RTX 2060 (6GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
Needs ~8 GB at Q4 GGUF, but only ~5 GB is usable on Nvidia GeForce RTX 2060 (6GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
The gap is about 3 GB: Wan 2.2 TI2V 5B needs roughly 8 GB and Nvidia GeForce RTX 2060 (6GB) leaves only about 5 GB usable for a model. The lightest tracked hardware that runs Wan 2.2 TI2V 5B is the Nvidia GeForce RTX 3060 (12GB) at 12 GB. See Wan 2.2 TI2V 5B on Nvidia GeForce RTX 3060 (12GB).
- Peak VRAM
- ~8 GB
- Usable on device
- ~5 GB
- Device memory
- 6 GB
- Quant
- Q4 GGUF
- Type
- video (DIT)
- Parameters
- 5B
- Peak VRAM
- ~8 GB at Q4 GGUF
- Resolution
- 1280×704 (720p)
- License
- Apache-2.0
- Memory
- 6 GB vram
- Usable for weights
- ~5 GB
- Power draw
- ~160 W
- Best runtime
- Ollama (CUDA) / llama.cpp CUDA
Run Wan 2.2 TI2V 5B on other hardware
FAQ
Can Nvidia GeForce RTX 2060 (6GB) run Wan 2.2 TI2V 5B?
Needs ~8 GB at Q4 GGUF, but only ~5 GB is usable on Nvidia GeForce RTX 2060 (6GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
How much VRAM does Wan 2.2 TI2V 5B need?
Nvidia GeForce RTX 2060 (6GB) does not have enough memory. At Q4 GGUF the realistic peak is ~8 GB of VRAM, versus ~24 GB with every component kept resident (no offload). With aggressive CPU offload it drops to ~5 GB, much slower.
What do I use to run Wan 2.2 TI2V 5B locally?
Wan 2.2 TI2V 5B runs in ComfyUI or Diffusers. It loads as a video diffusion checkpoint plus its text encoder and VAE, not a single chat command.
Sources
VRAM figures are sourced peak-usage anchors at the noted quant, validated 2026-09-14. See methodology.