2 citations · 3 across the 24 of their papers we have counts for
Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
Flatness Preserves Instruction Following in Vision-Language-Action Models
Haochen Zhang, Yonatan Bisk
Vision-language-action (VLA) models have the potential for open-world generalization by leveraging pretrained vision-language representations, yet downstream finetuning on limited…
cs.RO2026
Inductive Generalization for Robotic Manipulation
Annabella Macaluso, Haochen Zhang, Ishaan Masilamony +2
Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models learn structures that transfer…