2 papers
cs.RO2026
GeoHAT: Geometry-Adaptive Hybrid Action Transformer for Mobile Manipulation
Xiangyu Zhu, Renjun Wu, Luzhou Ge +2
Whole-body mobile manipulation requires coordinating mobile base and manipulator under shifting viewpoints, posing challenges in geometric perception and action generation. Current…
cs.CV2026
DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding
Luzhou Ge, Xiangyu Zhu, Jinyan Liu +1
Integrating open-vocabulary semantic information into dynamic 3D scene representations is essential for long-term embodied scene understanding. However, existing methods often suff…