works on

From the 1 of 14 linked papers with an AI index.

activity
20242026
collaborators

14 papers

cs.CV2026

InstanceSplat: Instance-Aware Feed-Forward 3D Gaussian Splatting for Scene Understanding

Minchao Jiang, Xiaoxuan Ma, Shunyu Jia +3

Feed-forward 3D Gaussian Splatting (3DGS) enables efficient and generalizable 3D reconstruction, but current feed-forward 3DGS methods for scene understanding remain largely catego…

cs.AI2026

Towards the Harness of Embodied Agents

Qi Wang, Tianyi Wang, Chengyang Li +6

The success of coding agents has established the harness as a paradigm: what an agent achieves depends not on the model alone, but on the infrastructure around it. We ask whether t…

cs.RO2026

Semantic Anchoring for Robotic Action Representations

Yuan Xu, Youheng Shi, Chengyang Li +2

The paper studies how fine‑tuning vision‑language‑action models for robots can degrade the semantic structure of their action representations, and proposes a plug‑and‑play anchorin…

cs.CV2026

Resonant Minds: Closed-Loop Social Avatars with Theory of Mind

Jianxu Shangguan, Jing Xu, Hang Ye +4

Creating lifelike digital humans with genuine social intelligence requires unifying cognitive reasoning and multimodal generation within a coherent framework. Current approaches tr…

cs.RO2026

Proprioceptive-visual correspondence enables self-other distinction in humanoid robots

Yurun Chen, Tianyuan Gao, Yizhong Ge +5

Distinguishing self from others is a prerequisite for social intelligence, yet humanoid robots that increasingly share workspaces with humans still lack this ability. Here we show…

cs.RO2026

GazeVLA: Learning Human Intention for Robotic Manipulation

Chengyang Li, Kaiyi Xiong, Yuan Xu +3

Embodied foundation models have achieved significant breakthroughs in robotic manipulation, yet they still depend heavily on large-scale robot demonstrations. Although recent works…