works on

From the 1 of 25 linked papers with an AI index.

collaborators

25 papers

cs.AI2026

WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN

Yuehao Huang, Yunzi Wu, Xiaotao Zhang +7

Recent vision-language navigation (VLN) systems increasingly adapt pretrained vision-language models (VLMs) into vision-language-action (VLA) policies that map egocentric observati…

cs.RO2026

KineBench: Benchmarking Embodied World Models via IDM-Free Kinematic Grounding

Zeyu Liu, Zhangzhe Zhu, Yang Zhang +3

Evaluating the physical consistency of embodied world models(EWMs) is a critical open challenge. While closed-loop evaluation via simulator rollouts offers a more faithful assessme…

cs.RO2026

EDAR: Learning Environment-Dependent Action Representations for Robotic Manipulation

Yuecheng Xu, Tong Yang, Jingkai Jia +3

The paper introduces EDAR, a method that learns action representations for robotic manipulation by linking control commands with the visual effects they cause in a given environmen…

cs.RO2026

KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills

Weiji Xie, Jinrui Han, Jiakun Zheng +6

Humanoid robots are promising to acquire various skills by imitating human behaviors. However, existing algorithms are only capable of tracking smooth, low-speed human motions, eve…

cs.RO2026

X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation

Dewei Wang, Xinmiao Wang, Chenyun Zhang +4

While recent advances have demonstrated strong performance in individual humanoid skills such as upright locomotion, fall recovery and whole-body coordination, learning a single po…

cs.RO2026

SpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and Reasoning

Yucheng Deng, Pingrui Lai, Xinhai Li +5

Vision-and-Language Navigation in continuous environments requires agents to understand the spatial structure of previously unseen environments in order to follow language instruct…