Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
PointAction: 3D Points as Universal Action Representations for Robot Control
Mutian Tong, Han Jiang, Qiao Feng +2
Video-Action Models (VAMs) leverage the broad visual dynamics captured by pre-trained video diffusion models, offering a promising path toward generalizable robot manipulation. How…
cs.RO2026
OmniGuide: Universal Guidance Fields for Enhancing Generalist Robot Policies
Yunzhou Song, Long Le, Yong-Hyun Park +7
Vision-language-action(VLA) models have shown great promise as generalist policies for a large range of relatively simple tasks. However, they demonstrate limited performance on mo…