Stable Video Diffusion Image to Video Model Card Stable Video Diffusion (SVD) Image to Video is a diffusion model that takes in a still image as a conditioning frame, and generates a video from it. Please note: For commercial use, please refer to https://stability.ai/license. Model Details Model Description (SVD) Image to Video is a latent diffusion model trained to generate short video clips from an image conditioning. This model was trained to generate 25 frames at resolution 576x1024 given a context frame of the same size, finetuned from [SVD Image to Video [14 frames]](https://huggingface.co/stabilityai/stable video diffusion img2vid). We also finetune the widely used f8 decoder for temporal consistency. For convenience, we additionally provide the model with the standard frame wise decoder here. Developed by: Stability AI Funded by: Stability AI Model type: Generative image to video model Finetuned from model: SVD Image to Video [14 frames] Model Sources For research purposes, we recommend our generative models Github repository (https://github.com/Stability AI/generative models), which implements the most popular diffusion frameworks (both training and inference). Repository:…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy