2 papers
cs.RO2026
RoboFollow: Unveiling the Instruction Following Mirage in Embodied Agents
Chang Guo, Yukun Xie, Bohan Tan +6
Modern embodied agents achieve impressive success rates, yet their actual instruction-following ability is far weaker than these numbers suggest. We trace this illusion to a struct…
cs.RO2026
Grounded Semantic Re-Binding for Robust Instruction Generalization in Vision-Language-Action Models
Zhaokai Yin, Zhipeng Zhang
Vision-Language-Action (VLA) models excel in robotic manipulation but suffer catastrophic performance drops when canonical instructions are simply paraphrased. Although this brittl…