From the 1 of 15 linked papers with an AI index.
9 papers · 1 filter
LIME: Learning Intent-aware Camera Motion from Egocentric Video
Boyang Sun, Jiajie Li, Yung-Hsu Yang +6
Autonomous robots often need to move their camera before they can act: to inspect an object, reveal an occluded region, or obtain a view that responds to a user's intent. While vis…
Hoi! - A Multimodal Dataset for Force-Grounded, Cross-View Articulated Manipulation
Tim Engelbracht, René Zurbrügg, Matteo Wohlrapp +5
We present a dataset for force-grounded, cross-view articulated manipulation that couples what is seen with what is done and what is felt during real human interaction. The dataset…
Loop Closure from Two Views: Revisiting PGO for Scalable Trajectory Estimation through Monocular Priors
Tian Yi Lim, Boyang Sun, Marc Pollefeys +1
(Visual) Simultaneous Localization and Mapping (SLAM) remains a fundamental challenge in enabling autonomous systems to navigate and understand large-scale environments. Traditiona…
ActLoc: Learning to Localize on the Move via Active Viewpoint Selection
Jiajie Li, Boyang Sun, Luca Di Giammarino +2
Reliable localization is critical for robot navigation, yet most existing systems implicitly assume that all viewing directions at a location are equally informative. In practice,…
FrontierNet: Learning Visual Cues to Explore
Boyang Sun, Hanzhi Chen, Stefan Leutenegger +3
Exploration of unknown environments is crucial for autonomous robots; it allows them to actively reason and decide on what new data to acquire for different tasks, such as mapping,…
Lost & Found: Tracking Changes from Egocentric Observations in 3D Dynamic Scene Graphs
Tjark Behrens, René Zurbrügg, Marc Pollefeys +2
Recent approaches have successfully focused on the segmentation of static reconstructions, thereby equipping downstream applications with semantic 3D understanding. However, the wo…