1 citations · 1 across the 3 of their papers we have counts for
Showing cs.ROShow all
3 papers · 1 filter
cs.RO2025
FrankenBot: Brain-Morphic Modular Orchestration for Robotic Manipulation with Vision-Language Models
Shiyi Wang, Wenbo Li, Yiteng Chen +2
Developing a general robot manipulation system capable of performing a wide range of tasks in complex, dynamic, and unstructured real-world environments has long been a challenging…
cs.RO2025
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
Yiteng Chen, Wenbo Li, Shiyi Wang +2
Building a general robotic manipulation system capable of performing a wide variety of tasks in real-world settings is a challenging task. Vision-Language Models (VLMs) have demons…
cs.RO2025★ 1 cited
AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making
Wenbo Li, Shiyi Wang, Yiteng Chen +2
Vision-Language Models (VLMs) encode knowledge and reasoning capabilities for robotic manipulation within high-dimensional representation spaces. However, current approaches often…