7 papers
PaMoSplat: Part-Aware Motion-Guided Gaussian Splatting for Dynamic Scene Reconstruction
Yinan Deng, Jianyu Dou, Jiahui Wang +3
Dynamic scene reconstruction represents a fundamental yet demanding challenge in computer vision and robotics. While recent progress in 3DGS-based methods has advanced dynamic scen…
See Once, Then Act: Vision-Language-Action Model with Task Learning from One-Shot Video Demonstrations
Guangyan Chen, Meiling Wang, Qi Shao +10
Developing robust and general-purpose manipulation policies represents a fundamental objective in robotics research. While Vision-Language-Action (VLA) models have demonstrated pro…
OmniMap: A General Mapping Framework Integrating Optics, Geometry, and Semantics
Yinan Deng, Yufeng Yue, Jianyu Dou +5
Robotic systems demand accurate and comprehensive 3D environment perception, requiring simultaneous capture of photo-realistic appearance (optical), precise layout shape (geometric…
OpenMulti: Open-Vocabulary Instance-Level Multi-Agent Distributed Implicit Mapping
Jianyu Dou, Yinan Deng, Jiahui Wang +3
Multi-agent distributed collaborative mapping provides comprehensive and efficient representations for robots. However, existing approaches lack instance-level awareness and semant…
OpenVox: Real-time Instance-level Open-vocabulary Probabilistic Voxel Representation
Yinan Deng, Bicheng Yao, Yihang Tang +2
In recent years, vision-language models (VLMs) have advanced open-vocabulary mapping, enabling mobile robots to simultaneously achieve environmental reconstruction and high-level s…
OpenIN: Open-Vocabulary Instance-Oriented Navigation in Dynamic Domestic Environments
Yujie Tang, Meiling Wang, Yinan Deng +3
In daily domestic settings, frequently used objects like cups often have unfixed positions and multiple instances within the same category, and their carriers frequently change as…