3 papers
cs.RO2025
FMimic: Foundation Models are Fine-grained Action Learners from Human Videos
Guangyan Chen, Meiling Wang, Te Cui +8
Visual imitation learning (VIL) provides an efficient and intuitive strategy for robotic systems to acquire novel skills. Recent advancements in foundation models, particularly Vis…
cs.RO2025
Human Demonstrations are Generalizable Knowledge for Robots
Te Cui, Tianxing Zhou, Zicai Peng +6
Learning from human demonstrations is an emerging trend for designing intelligent robotic systems. However, previous methods typically regard videos as instructions, simply dividin…
cs.RO2024
VLMimic: Vision Language Models are Visual Imitation Learner for Fine-grained Actions
Guanyan Chen, Meiling Wang, Te Cui +9
Visual imitation learning (VIL) provides an efficient and intuitive strategy for robotic systems to acquire novel skills. Recent advancements in Vision Language Models (VLMs) have…