1 paper
Kaichen Zhou, Yuhan Wang, Grace Chen +5
Recent 3D feed-forward models, such as the Visual Geometry Grounded Transformer (VGGT), have shown strong capability in inferring 3D attributes of static scenes. However, since the…