VGGT: Visual Geometry Grounded Transformer Meta AI Research ; University of Oxford, VGG Jianyuan Wang, Minghao Chen, Nikita Karaev, Andrea Vedaldi, Christian Rupprecht, David Novotny Overview Visual Geometry Grounded Transformer (VGGT, CVPR 2025) is a feed forward neural network that directly infers all key 3D attributes of a scene, including extrinsic and intrinsic camera parameters, point maps, depth maps, and 3D point tracks, from one, a few, or hundreds of its views, within seconds . Quick Start Please refer to our Github Repo Citation If you find our repository useful, please consider giving it a star ⭐ and citing our paper in your work:
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy