Wan2.2 S2V 14B: Audio Driven Cinematic Video Generation This repository features the Wan2.2 S2V 14B model, designed for audio driven cinematic video generation. It was introduced in the paper: Wan S2V: Audio Driven Cinematic Video Generation 💜 Wan Homepage    |    🖥️ GitHub      🤗 Hugging Face Organization      🤖 ModelScope Organization       📑 Wan S2V Paper       📑 Wan2.2 Base Paper    🌐 Project Page       📑 Blog       💬 Discord    📕 使用指南(中文)       📘 User Guide(English)      💬 WeChat(微信)    Abstract (Wan S2V Paper) Current state of the art (SOTA) methods for audio driven character animation demonstrate promising performance for scenarios primarily involving speech and singing. However, they often fall short in more complex film and television productions, which demand sophisticated elements such as nuanced character interactions, realistic body movements, and dynamic camera work. To address this long standing challenge of achieving film level character animation, we propose an audio driven model, which w…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy