FastVideo FastWan2.2 TI2V 5B FullAttn Diffusers Model FastVideo Team   HF Paper (VSA) arXiv Paper (VSA) Github Project Page Online Demo You can try our models here! Introduction We're excited to introduce the FastWan2.2 series —a new line of models finetuned with our novel Sparse distill strategy. This approach jointly integrates DMD and VSA in a single training process, combining the benefits of both distillation to shorten diffusion steps and sparse attention to reduce attention computations, enabling even faster video generation. FastWan2.2 TI2V 5B Full Diffusers is built upon Wan AI/Wan2.2 TI2V 5B Diffusers. It supports efficient 3 step inference and produces high quality videos at 121×704×1280 resolution. For training, we used simulated forward for the generator model, making the process data free. The current FastWan2.2 TI2V 5B Full Diffusers model is trained using only DMD . Model Overview 3 step inference is supported. Our model is trained on 121×704×1280 resolution, but it supports generating videos with any resolution .(quality may degrade) Finetuning and inference scripts are available in the FastVideo repository: 1 Node/GPU debugging finetuning script Slurm trainin…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy