2 papers
cs.RO2026
Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation
He Kong, Zengjue Chen, Qi Wang +6
Vision-language-action (VLA) models have demonstrated remarkable capabilities in robotic manipulation by leveraging pretrained vision-language models. However, existing post-traini…
cs.CV2026
LAWM-3D: Learning 3D-Aware Latent Actions from Human Videos for Generalizable Robot World Models
Jiarui Yang, Jiale Zhange, Jiawei Li +5
World models enable agents to perform forward rollout and planning without real-world interaction. However, their application in open-world embodied intelligence remains limited by…