From the 1 of 6 linked papers with an AI index.
2 citations · 2 across the 5 of their papers we have counts for
6 papers
4D-WAM: Infusing Spatiotemporal Awareness into World Action Models through Trajectory Fields
Lishan Yang, Wenxuan Song, Xi Wang +14
Building on recent advances in world models, World Action Models (WAMs) jointly model video prediction and action generation. However, they typically represent videos in 2D pixel s…
DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation
Junfeng Li, Junjie He, Zhide Zhong +12
Vision-Language-Action (VLA) models have become a powerful paradigm for robot manipulation, but training a single generalist policy for heterogeneous robot embodiments remains an o…
Step-Level Preference Learning for Generative Agents in Social Simulations
Wenchang Gao, Pingyue Sheng, Lanlan Qiu +7
The paper presents an interactive interface to collect step‑level human preference data for generative agents, creates a 57K annotation dataset, and shows that training LLMs with t…
Data Scaling Laws in Imitation Learning for Robotic Manipulation
Fanqi Lin, Yingdong Hu, Pingyue Sheng +3
Data scaling has revolutionized fields like natural language processing and computer vision, providing models with remarkable generalization capabilities. In this paper, we investi…
Prior Reinforce: Goal-Conditioned Dynamic Manipulation with Limited Trials
Yihang Hu, Pingyue Sheng, Yuyang Liu +2
Embodied robots have achieved strong performance in many real-world manipulation tasks, yet agile dynamic manipulation remains challenging due to high sensitivity to motion paramet…
Can Large Language Models Reinvent Foundational Algorithms?
Jian Zhao, Haoren Luo, Yu Wang +3
LLMs have shown strong potential to advance scientific discovery. Whether they possess the capacity for foundational innovation, however, remains an open question. In this work, we…