Quantized GGUFs of LongCat Video Avatar 1.5 for ComfyUI + WanVideoWrapper Original model Link: https://huggingface.co/meituan longcat/LongCat Video Avatar 1.5 Watch us at Youtube: @VantageWithAI LongCat Video Avatar 1.5 π Model Introduction We are excited to announce the release of LongCat Video Avatar 1.5, an upgraded open source framework that prioritizes extreme empirical optimization and production readiness for audio driven human video generation. Built upon the LongCat Video foundation model, v1.5 delivers highly stable, commercial grade avatar video synthesis supporting native tasks including Audio Text to Video (AT2V), Audio Text Image to Video (ATI2V), and Video Continuation, with seamless compatibility for both single stream and multi stream audio inputs. Key Features π Upgraded Audio Encoder (Whisper Large): : Replaces Wav2Vec2 with Whisper Large, yielding significantly smoother and more natural lip dynamics. π Production Ready Stability : Achieves accurate lip synchronization, full body temporal stability, and robust long video generation with strict identity consistency. π Stylized Domain Generalization : Robustly generalizes to anime, animals, and complex real worβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy