video model · wan · Windows
Can I run Wan 2.2 TI2V 5B on Nvidia GeForce GTX 1070 (8GB)?
Needs ~8 GB at Q4 GGUF, but only ~7 GB is usable on Nvidia GeForce GTX 1070 (8GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
Needs ~8 GB at Q4 GGUF, but only ~7 GB is usable on Nvidia GeForce GTX 1070 (8GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
The gap is about 1 GB: Wan 2.2 TI2V 5B needs roughly 8 GB and Nvidia GeForce GTX 1070 (8GB) leaves only about 7 GB usable for a model. The lightest tracked hardware that runs Wan 2.2 TI2V 5B is the Nvidia GeForce RTX 3060 (12GB) at 12 GB. See Wan 2.2 TI2V 5B on Nvidia GeForce RTX 3060 (12GB).
- Peak VRAM
- ~8 GB
- Usable on device
- ~7 GB
- Device memory
- 8 GB
- Quant
- Q4 GGUF
- Type
- video (DIT)
- Parameters
- 5B
- Peak VRAM
- ~8 GB at Q4 GGUF
- Resolution
- 1280×704 (720p)
- License
- Apache-2.0
- Memory
- 8 GB vram
- Usable for weights
- ~7 GB
- Power draw
- ~150 W
- Best runtime
- llama.cpp CUDA (Pascal, no tensor cores)
What you can run instead
Run Wan 2.2 TI2V 5B on other hardware
FAQ
Can Nvidia GeForce GTX 1070 (8GB) run Wan 2.2 TI2V 5B?
Needs ~8 GB at Q4 GGUF, but only ~7 GB is usable on Nvidia GeForce GTX 1070 (8GB). With aggressive CPU offload it can run on as little as ~5 GB, much slower.
How much VRAM does Wan 2.2 TI2V 5B need?
Nvidia GeForce GTX 1070 (8GB) does not have enough memory. At Q4 GGUF the realistic peak is ~8 GB of VRAM, versus ~24 GB with every component kept resident (no offload). With aggressive CPU offload it drops to ~5 GB, much slower.
What do I use to run Wan 2.2 TI2V 5B locally?
Wan 2.2 TI2V 5B runs in ComfyUI or Diffusers. It loads as a video diffusion checkpoint plus its text encoder and VAE, not a single chat command.
Sources
VRAM figures are sourced peak-usage anchors at the noted quant, validated 2026-09-14. See methodology.