This repository contains the data presented in Video R1: Reinforcing Video Reasoning in MLLMs. Code: https://github.com/tulerfeng/Video R1 Video data folder: CLEVRER, LLaVA Video 178K, NeXT QA, PerceptionTest, STAR Image data folder: Chart, General, Knowledge, Math, OCR, Spatial Video R1 COT 165k.json is for SFT cold start, and Video R1 260k.json is for RL training. Data Format in Video R1 COT 165k: { "problem id": 2, "problem": "What appears on the screen in Russian during the… See the full description on the dataset page: https://huggingface.co/datasets/Huangzx1023/Video Training.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy