3 papers
cs.RO2025
Robotic Assistant: Completing Collaborative Tasks with Dexterous Vision-Language-Action Models
Boshi An, Chenyu Yang, Robert Katzschmann
We adapt a pre-trained Vision-Language-Action (VLA) model (Open-VLA) for dexterous human-robot collaboration with minimal language prompting. Our approach adds (i) FiLM conditionin…
cs.RO2025
Arnold: A multi-task, multi-embodiment muscle transformer policy
Alberto Silvio Chiappa, Boshi An, Merkourios Simos +2
Controlling high-dimensional and nonlinear musculoskeletal models of the human body is a foundational scientific challenge. Recent machine learning breakthroughs have heralded in-s…
cs.RO2025
RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning
Haoran Geng, Feishi Wang, Songlin Wei +34
Data scaling and standardized evaluation benchmarks have driven significant advances in natural language processing and computer vision. However, robotics faces unique challenges i…