3 citations · 3 across the 9 of their papers we have counts for
9 papers
GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes
Ermanno Bartoli, Buwei He, Dennis Rotondi +6
Robots operating in human environments need memories that capture not only what objects exist and where, but also how people use them over time and how individual interactions comp…
FLAT: Feedforward Latent Triangle Splatting for Geometrically Accurate Scene Generation
Orest Kupyn, Goutam Bhat, Philipp Henzler +3
Generating explorable 3D scenes from a single image requires strong generative priors and accurate geometric representations suitable for downstream use. Current video diffusion mo…
Prompting Diffusion Models for Zero-Shot Instance Segmentation
Irem Zeynep Alagöz, Nils Morbitzer, Andrea Ramazzina +3
Several disruptive research directions have recently emerged in computer vision, including foundation models achieving previously unseen zero-shot performance in scene understandin…
The Art of Interrogation: Consistency Amplifies Factuality in Spatial Reasoning
Theo Uscidda, Marta Tintore Gazulla, Maks Ovsjanikov +2
Current Large Reasoning Models (LRMs) exhibit remarkable general capabilities but significantly underperform in spatial reasoning tasks. Existing approaches treat this gap as a kno…
3D Scene Graphs: Open Challenges and Future Directions
Dennis Rotondi, Francesco Argenziano, Sebastian Koch +10
3D Scene Graphs (3DSGs) have emerged as a powerful representation for spatial AI by combining geometric grounding with semantic and relational abstractions of the environment. Thei…
DynaTok: Token-Based 4D Reconstruction from Partial Point Clouds
Weirong Chen, Keisuke Tateno, Hidenobu Matsuki +3
We address 4D reconstruction from partial point cloud sequences, where depth-sensor observations are incomplete, unordered, and lack explicit temporal correspondences. This geometr…