We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy
VidChapters Dataset for Chapter-Llama This repository contains the dataset used in the paper "Chapter-Llama: Efficient Chaptering in Hour-Long Videos with LLMs" (CVPR 2025). Overview VidChapters-7M is a large-scale dataset for video chaptering, containing: 817k videos with ASR data (20GB) Captions extracted from videos using various sampling strategies Chapter annotations with timestamps and titles Data Structure The dataset is organized as… See the full description on the dataset page: https://huggingface.co/datasets/lucas-ventura/chapter-llama.
No dataset card provided yet.
Mirrored from an external registry.
Last synced 6/14/2026
Preview not yet available for this dataset