From the 1 of 25 linked papers with an AI index.
25 papers
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
Yuehao Huang, Yunzi Wu, Xiaotao Zhang +7
Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into vision-language-action (VLA) policies that map egocentric observati…
KineBench: Benchmarking Embodied World Models via IDM-Free Kinematic Grounding
Zeyu Liu, Zhangzhe Zhu, Yang Zhang +3
Evaluating the physical consistency of embodied world models(EWMs) is a critical open challenge. While closed-loop evaluation via simulator rollouts offers a more faithful assessme…
EDAR: Learning Environment-Dependent Action Representations for Robotic Manipulation
Yuecheng Xu, Tong Yang, Jingkai Jia +3
The paper introduces EDAR, a method that learns action representations for robotic manipulation by linking control commands with the visual effects they cause in a given environmen…
KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills
Weiji Xie, Jinrui Han, Jiakun Zheng +6
Humanoid robots are promising to acquire various skills by imitating human behaviors. However, existing algorithms are only capable of tracking smooth, low-speed human motions, eve…
X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation
Dewei Wang, Xinmiao Wang, Chenyun Zhang +4
While recent advances have demonstrated strong performance in individual humanoid skills such as upright locomotion, fall recovery and whole-body coordination, learning a single po…
SpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and Reasoning
Yucheng Deng, Pingrui Lai, Xinhai Li +5
Vision-and-Language Navigation in continuous environments requires agents to understand the spatial structure of previously unseen environments in order to follow language instruct…