1 paper · 1 filter
Chenyi Wang, Xinkai Wang, Bokai Lin +4
Action labels tell a vision-language-action (VLA) policy which robot commands to imitate, but not how those commands change the 3D world. The aligned demonstration clip contains th…