2 papers
cs.RO2026
SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
Anna Gelencsér-Horváth, Gergely Dinya, Dorka Boglárka Erős +3
We present SceneVGGT, a spatio-temporal 3D scene understanding framework that combines SLAM with semantic mapping for autonomous and assistive navigation. Built on VGGT, our method…
cs.CV2025
Building temporally coherent 3D maps with VGGT for memory-efficient Semantic SLAM
Gergely Dinya, Péter Halász, András Lőrincz +2
We present a fast, spatio-temporal scene understanding framework based on Visual Geometry Grounded Transformer (VGGT). The proposed pipeline is designed to enable efficient, close…