AVGen Bench Generated Videos Data Card Overview This data card describes the generated audio video outputs stored directly in the repository root by model directory. The collection is intended for benchmarking and qualitative/quantitative evaluation of text to audio video (T2AV) systems. It was presented in the paper AVGen Bench: A Task Driven Benchmark for Multi Granular Evaluation of Text to Audio Video Generation. It is not a training dataset. Each item is a model generated video produced from a prompt defined in prompts/ .json . For Hugging Face Hub compatibility, the repository includes a root level metadata.parquet file so the Dataset Viewer can expose each video as a structured row with prompt metadata instead of treating the repo as an unindexed file dump. The relative video path is stored as a plain string column ( video path ) rather than a media typed file name column, which avoids current Dataset Viewer post processing failures on video rows. Sample Usage As described in the GitHub repository, you can generate videos from the benchmark prompts using the following command: What This Dataset Contains The dataset is organized by: 1. Model directory 2. Video category 3. Gen…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy