4 citations · 4 across the 3 of their papers we have counts for
7 papers · 1 filter
Virtual Community: An Open World for Humans, Robots, and Society
Qinhong Zhou, Hongxin Zhang, Xiangye Lin +16
The rapid progress in AI and Robotics may lead to a profound societal transformation, as humans and robots begin to coexist within shared communities, introducing both opportunitie…
LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences
Hongyan Zhi, Peihao Chen, Junyan Li +6
Research on 3D Vision-Language Models (3D-VLMs) is gaining increasing attention, which is crucial for developing embodied AI within 3D scenes, such as visual navigation and embodie…
CoNav: A Benchmark for Human-Centered Collaborative Navigation
Changhao Li, Xinyu Sun, Peihao Chen +6
Human-robot collaboration, in which the robot intelligently assists the human with the upcoming task, is an appealing objective. To achieve this goal, the agent needs to be equippe…
SKDF: A Simple Knowledge Distillation Framework for Distilling Open-Vocabulary Knowledge to Open-world Object Detector
Shuailei Ma, Yuefeng Wang, Ying Wei +4
In this paper, we attempt to specialize the VLM model for OWOD tasks by distilling its open-world knowledge into a language-agnostic detector. Surprisingly, we observe that the com…
Contrastive Vision-Language Alignment Makes Efficient Instruction Learner
Lizhao Liu, Xinyu Sun, Tianhang Xiang +3
We study the task of extending the large language model (LLM) into a vision-language instruction-following model. This task is crucial but challenging since the LLM is trained on t…
FGPrompt: Fine-grained Goal Prompting for Image-goal Navigation
Xinyu Sun, Peihao Chen, Jugang Fan +3
Learning to navigate to an image-specified goal is an important but challenging task for autonomous systems. The agent is required to reason the goal location from where a picture…