2 papers
cs.RO2026
Affordance-Aware Interactive Decision-Making and Execution for Ambiguous Instructions
Hengxuan Xu, Fengbo Lan, Zhixin Zhao +5
Enabling robots to explore and act in unfamiliar environments under ambiguous human instructions by interactively identifying task-relevant objects (e.g., identifying cups or bever…
cs.RO2025
RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation
Sheng Wang
As robotic technologies advancing towards more complex multimodal interactions and manipulation tasks, the integration of advanced Vision-Language Models (VLMs) has become a key dr…