DL3DV Depth DA3 Aligned Per frame depth annotations for the DL3DV dataset, produced by Depth Anything 3 (DA3) and then aligned to each scene's sparse depth from the original DL3DV reconstruction. We use this refined dataset for 3D world generation and reconstruction in our Puffin World . Sample Videos Each clip is a 1×3 comparison — RGB Original Depth Our Aligned Depth — with depth rendered by vision banana representation. It shows how the DA3 aligned depth (right) densifies and cleans up the original sparse depth (middle). Directory structure The archive mirrors the source DL3DV layout — one .zip per scene, grouped into bucket folders 1K – 7K ( / .zip ). The scene hashes match DL3DV and DL3DV Absolute Camera, so depth pairs 1:1 with the source frames / absolute camera annotations. Each .zip unpacks to: Each frame NNNNN.npy is a float32 depth map — np.load(...) returns an array of shape (H, W) (e.g. (536, 954) ), one per source frame, indices matching the DL3DV frames. How the depth was produced Predicted with Depth Anything 3 (DA3). Aligned to the sparse depth of the original DL3DV dataset (per scene alignment against the sparse reconstruction), so each scene's DA3 depth is brough…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy