2 papers
cs.LG2026
: A Vision-Language-Action Flow Model for General Robot Control
Kevin Black, Noah Brown, Danny Driess +21
Robot learning holds tremendous promise to unlock the full potential of flexible, general, and dexterous robot systems, as well as to address some of the deepest questions in artif…
cs.LG2025
: a Vision-Language-Action Model with Open-World Generalization
Physical Intelligence, Kevin Black, Noah Brown +33
In order for robots to be useful, they must perform practically relevant tasks in the real world, outside of the lab. While vision-language-action (VLA) models have demonstrated im…