$\pi^3$: Permutation Equivariant Visual Geometry Learning This repository contains the weights for Pi3X , an enhanced version of the $\pi^3$ model introduced in the paper $\pi^3$: Permutation Equivariant Visual Geometry Learning. $\pi^3$ is a feed forward neural network for visual geometry reconstruction that eliminates the need for a fixed reference view. It employs a fully permutation equivariant architecture to predict affine invariant camera poses and scale invariant local point maps from an unordered set of images, making it robust to input ordering and achieving state of the art performance. Project Page: yyfz.github.io/pi3/ GitHub Repository: github.com/yyfz/Pi3 Demo: Hugging Face Space Pi3X Engineering Update Pi3X is an enhanced version focusing on flexibility and reconstruction quality: Smoother Reconstruction: Uses a Convolutional Head to reduce grid like artifacts. Flexible Conditioning: Supports optional injection of camera poses, intrinsics, and depth. Reliable Confidence: Predicts continuous quality levels for better noise filtering. Metric Scale: Supports approximate metric scale reconstruction. Sample Usage To use this model, you need to clone the official repositor…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy