VGen VGen is an open source video synthesis codebase developed by the Tongyi Lab of Alibaba Group, featuring state of the art video generative models. This repository includes implementations of the following methods: I2VGen xl: High quality image to video synthesis via cascaded diffusion models VideoComposer: Compositional Video Synthesis with Motion Controllability Hierarchical Spatio temporal Decoupling for Text to Video Generation A Recipe for Scaling up Text to Video Generation with Text free Videos InstructVideo: Instructing Video Diffusion Models with Human Feedback DreamVideo: Composing Your Dream Videos with Customized Subject and Motion VideoLCM: Video Latent Consistency Model Modelscope text to video technical report VGen can produce high quality videos from the input text, images, desired motion, desired subjects, and even the feedback signals provided. It also offers a variety of commonly used video generation tools such as visualization, sampling, training, inference, join training using images and videos, acceleration, and more. 🔥News!!! [2023.12] We release the high efficiency video generation method VideoLCM [2023.12] We release the code and model of I2VGen XL and…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy