1 paper · 1 filter
Yongsung Kim, Wooseok Song, Jaihyun Lew +3
Visual Geometry Grounded Transformer (VGGT) has advanced 3D vision, yet its global attention layers suffer from quadratic computational costs that hinder scalability. Several spars…