🧠 Project SLOB: Spontaneous Lifestyle & Observational Behaviors Dataset 📌 Abstract Welcome to the primary data ingestion node for Project SLOB . This repository hosts a massive, high fidelity multimodal dataset designed to train next generation Artificial Intelligence in recognizing, analyzing, and predicting spontaneous human behaviors in unconstrained, real world video streams. This node is strictly used for the Spatial Temporal Audio Visual Synchronization (ST AVS) phase of the SLOB architecture. 📂 Dataset Architecture & Modalities This repository operates as a dynamic, rolling pipeline. You will encounter various file types which are outputs of our multi pass processing nodes: Raw Video Streams ( video.mp4 ): High framerate visual data used for spatial tracking and behavioral bounding box generation. Synthetic Audio Injections ( audio vi tts.mp3 ): AI generated audio arrays used to test the model's robustness against audio visual desynchronization and cross lingual hallucination. Time Warped Subtitles ( VI STRETCHED.srt ): Transcripts that have been deliberately dilated (e.g., 0.85x speed factor) to train the AI's temporal alignment algorithms. Rendered Composites ( FINAL .m…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy