1 paper
Jinhao You, Shuo Lyu, Zhuohang Lyu +5
Visual Geometry Grounded Transformer (VGGT) recovers dense 3D scene structure from multi-view images in one forward pass, but quadratic cross-frame attention limits its scalability…