5 papers
GESTO: Human-Centric Spatio-Temporal Memory for Reasoning in Dynamic Scenes
Ermanno Bartoli, Buwei He, Dennis Rotondi +6
Robots operating in human environments need memories that capture not only what objects exist and where, but also how people use them over time and how individual interactions comp…
Learning to Localize Reference Trajectories in Image-Space for Visual Navigation
Finn Lukas Busch, Matti Vahs, Quantao Yang +4
We present LoTIS, a model for visual navigation that provides robot-agnostic image-space guidance by localizing a reference RGB trajectory in the robot's current view, without requ…
PRIX: Learning to Plan from Raw Pixels for End-to-End Autonomous Driving
Maciej K. Wozniak, Lianhang Liu, Yixi Cai +1
While end-to-end autonomous driving models show promising results, their practical deployment is often hindered by large model sizes, a reliance on expensive LiDAR sensors and comp…
DeltaFlow: An Efficient Multi-frame Scene Flow Estimation Method
Qingwen Zhang, Xiaomeng Zhu, Yushan Zhang +3
Previous dominant methods for scene flow estimation focus mainly on input from two consecutive frames, neglecting valuable information in the temporal domain. While recent trends s…
DoGFlow: Self-Supervised LiDAR Scene Flow via Cross-Modal Doppler Guidance
Ajinkya Khoche, Qingwen Zhang, Yixi Cai +2
Accurate 3D scene flow estimation is critical for autonomous systems to navigate dynamic environments safely, but creating the necessary large-scale, manually annotated datasets re…