3 papers
cs.RO2026
Symskill: Symbol and Skill Co-Invention for Data-Efficient and Reactive Long-Horizon Manipulation
Yifei Simon Shao, Yuchen Zheng, Sunan Sun +3
Multi-step manipulation in dynamic environments remains challenging. Imitation learning (IL) is reactive but lacks compositional generalization, since monolithic policies do not de…
cs.RO2025
VLMgineer: Vision Language Models as Robotic Toolsmiths
George Jiayuan Gao, Tianyu Li, Junyao Shi +4
Tool design and use reflect the ability to understand and manipulate the physical world through creativity, planning, and foresight. As such, these capabilities are often regarded…
cs.RO2024
Don't Yell at Your Robot: Physical Correction as the Collaborative Interface for Language Model Powered Robots
Chuye Zhang, Yifei Simon Shao, Harshil Parekh +4
We present a novel approach for enhancing human-robot collaboration using physical interactions for real-time error correction of large language model (LLM) powered robots. Unlike…