Quantized GGUF version of Bernini R for ComfyUI. Original model link: https://huggingface.co/ByteDance/Bernini R Watch us on Youtube: @VantageWithAI Latent Semantic Planning for Video Diffusion Chenchen Liu \ , Junyi Chen \ , Lei Li \ , Lu Chi \ ,§ , Mingzhen Sun \ , Zhuoying Li \ , Yi Fu, Ruoyu Guo, Yiheng Wu, Ge Bai, Zehuan Yuan ✉ \ Equal contribution ✉ Corresponding author § Project lead 🎉 News [2026 06 01] We open sourced the inference code and model weights of the Bernini Renderer ( Bernini R ). [2026 05 22] We released our paper Bernini: Latent Semantic Planning for Video Diffusion. ✨ Highlights Bernini is a unified framework for video generation and editing that combines an MLLM based semantic planner with a DiT based renderer. On video editing, Bernini reaches the first tier among leading closed source commercial models. The leaderboard below comes from our self built arena platform, where human annotators blindly vote on paired edits and the votes are aggregated into a Bradley Terry score and a pairwise win rate matrix. 📑 Citation If you use Bernini in your research, please cite: 🙏 Acknowledgements Bernini builds on several outstanding open sourc…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy