Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
Fine-Tuning VLAs with Self-Demonstrated Generative Control for Multi-Task Manipulation
Prachi Garg, Steve Xing, Prahit Yaugand +2
State-of-the-art vision-language-action (VLA) models such as exhibit strong semantic understanding, instruction following and task behavior. However, when deployed on new…
cs.RO2024
Push Past Green: Learning to Look Behind Plant Foliage by Moving It
Xiaoyu Zhang, Saurabh Gupta
Autonomous agriculture applications (e.g., inspection, phenotyping, plucking fruits) require manipulating the plant foliage to look behind the leaves and the branches. Partial visi…