From the 1 of 14 linked papers with an AI index.
14 papers
InstanceSplat: Instance-Aware Feed-Forward 3D Gaussian Splatting for Scene Understanding
Minchao Jiang, Xiaoxuan Ma, Shunyu Jia +3
Feed-forward 3D Gaussian Splatting (3DGS) enables efficient and generalizable 3D reconstruction, but current feed-forward 3DGS methods for scene understanding remain largely catego…
Towards the Harness of Embodied Agents
Qi Wang, Tianyi Wang, Chengyang Li +6
The success of coding agents has established the harness as a paradigm: what an agent achieves depends not on the model alone, but on the infrastructure around it. We ask whether t…
Semantic Anchoring for Robotic Action Representations
Yuan Xu, Youheng Shi, Chengyang Li +2
The paper studies how fine‑tuning vision‑language‑action models for robots can degrade the semantic structure of their action representations, and proposes a plug‑and‑play anchorin…
Resonant Minds: Closed-Loop Social Avatars with Theory of Mind
Jianxu Shangguan, Jing Xu, Hang Ye +4
Creating lifelike digital humans with genuine social intelligence requires unifying cognitive reasoning and multimodal generation within a coherent framework. Current approaches tr…
Proprioceptive-visual correspondence enables self-other distinction in humanoid robots
Yurun Chen, Tianyuan Gao, Yizhong Ge +5
Distinguishing self from others is a prerequisite for social intelligence, yet humanoid robots that increasingly share workspaces with humans still lack this ability. Here we show…
GazeVLA: Learning Human Intention for Robotic Manipulation
Chengyang Li, Kaiyi Xiong, Yuan Xu +3
Embodied foundation models have achieved significant breakthroughs in robotic manipulation, yet they still depend heavily on large-scale robot demonstrations. Although recent works…