🚗 OmniTraffic: A Large scale Multi view Spatiotemporal Benchmark for Traffic Understanding 🤗 Hugging Face Dataset • 📃 Paper (Coming Soon) • 💻 GitHub Repo • 🏆 Leaderboard 📌 Benchmark Summary OmniTraffic is a comprehensive evaluation benchmark designed to test the multi view spatiotemporal reasoning and Bird's Eye View (BEV) perception capabilities of multimodal large language models (MLLMs) and autonomous driving systems. While the complete OmniTraffic dataset ecosystem contains an underlying pool of over 8 million generated VQA samples , this repository specifically hosts the OmniTraffic Gold Standard Benchmark . It consists of 3,200 highly curated VQA pairs that were systematically sampled from the massive 8M pool and rigorously validated by human experts. This curated subset ensures logically sound reasoning chains and eliminates low quality noise, serving as a reliable metric for deep traffic logic evaluation. Figure 1: Distribution of the 3,200 curated VQA pairs across the three evaluation levels, task categories, and required capabilities. 🌟 Key Features Gold Standard Quality : 3,200 human validated QA pairs sampled from an 8M+ massive data pool. Spatiotemporal Reasonin…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy