ConsIDVid Summary This dataset is released with our paper: ConsID Gen: View Consistent and Identity Preserving Image to Video Generation Mingyang Wu, Ashirbad Mishra, Soumik Dey, Shuo Xing, Naveen Ravipati, Hansi Wu, Binbin Li, Zhengzhong Tu (2026) Accepted by CVPR 2026 This release contains processed videos and hierarchical captions used in ConsID Gen experiments across: public synthesis (MVImgNet) public (CO3D, OmniObject3D, Objectron) The dataset is designed for image to video generation research with an emphasis on identity preservation and view consistency. Paper and Project Paper (arXiv): https://arxiv.org/abs/2602.10113 Project page: https://myangwu.github.io/ConsID Gen/ Directory Current Release Notes Current snapshot includes the folders listed above. More data is currently under internal review and will be released after approval. Download Use the Hugging Face CLI: Or download specific subsets: Usage Captions are stored as .txt files in hierarchy video caption/{source}/{s1,s2} . Videos are stored as .mp4 files in processed videos/{source} . Example: Citation
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy