10 citations · 12 across the 7 of their papers we have counts for
14 papers
Pandora: Articulated 3D Scene Graphs from Egocentric Vision
Alan Yu, Yun Chang, Christopher Xie +1
Robotic mapping systems typically approach building metric-semantic scene representations from the robot's own sensors and cameras. However, these "first person" maps inherit the r…
VGGT-SLAM 2.0: Real-time Dense Feed-forward Scene Reconstruction
Dominic Maggio, Luca Carlone
We present VGGT-SLAM 2.0, a real-time RGB feed-forward SLAM system which substantially improves upon VGGT-SLAM for incrementally aligning submaps created from VGGT. Firstly, we rem…
Towards Zero-Shot Point Cloud Registration Across Diverse Scales, Scenes, and Sensor Setups
Hyungtae Lim, Minkyun Seo, Luca Carlone +1
Some deep learning-based point cloud registration methods struggle with zero-shot generalization, often requiring dataset-specific hyperparameter tuning or retraining for new envir…
Describe Anything Anywhere At Any Moment
Nicolas Gorlo, Lukas Schmid, Luca Carlone
Computer vision and robotics applications ranging from augmented reality to robot autonomy in large-scale environments require spatio-temporal memory frameworks that capture both g…
Advancing AI Challenges for the United States Department of the Air Force
Christian Prothmann, Vijay Gadepally, Jeremy Kepner +35
The DAF-MIT AI Accelerator is a collaboration between the United States Department of the Air Force (DAF) and the Massachusetts Institute of Technology (MIT). This program pioneers…
Structured Interfaces for Automated Reasoning with 3D Scene Graphs
Aaron Ray, Jacob Arkin, Harel Biggie +3
In order to provide a robot with the ability to understand and react to a user's natural language inputs, the natural language must be connected to the robot's underlying represent…