8 papers
Scaling Self-Play for End-to-End Driving
Luke Rowe, Roger Girgis, Rodrigue de Schaetzen +6
End-to-end autonomous driving models are typically trained on offline human-demonstration datasets that provide limited state coverage and often no closed-loop feedback, making the…
FlowMaps: Modeling Long-Term Multimodal Object Dynamics with Flow Matching
Francesco Argenziano, Miguel Saavedra-Ruiz, Sacha Morin +3
Joint spatial and temporal understanding of 3D scenes is a crucial requirement for robots deployed in everyday household environments. Such agents must not only comprehend and navi…
3D Scene Graphs: Open Challenges and Future Directions
Dennis Rotondi, Francesco Argenziano, Sebastian Koch +10
3D Scene Graphs (3DSGs) have emerged as a powerful representation for spatial AI by combining geometric grounding with semantic and relational abstractions of the environment. Thei…
PerceptTwin: Semantic Scene Reconstruction for Iterative LLM Planning and Verification
Charlie Gauthier, Sacha Morin, Liam Paull
Simulation environments are useful for both robot policy learning and planning verification and validation. Traditionally, the process of creating a simulation was onerous. Creatin…
OpenLex3D: A Tiered Evaluation Benchmark for Open-Vocabulary 3D Scene Representations
Christina Kassab, Sacha Morin, Martin Büchner +5
3D scene understanding has been transformed by open-vocabulary language models that enable interaction via natural language. However, at present the evaluation of these representat…
Agentic Scene Policies: Unifying Space, Semantics, and Affordances for Robot Action
Sacha Morin, Kumaraditya Gupta, Mahtab Sandhu +4
Executing open-ended natural language queries is a core problem in robotics. While recent advances in imitation learning and vision-language-actions models (VLAs) have enabled prom…