9 citations · 9 across the 2 of their papers we have counts for
Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
Gripper-aware Vision Language Action Models
Hanyi Zhang, Zihong Luo, Tianyu Li +16
Vision language action models (VLAs) have advanced general purpose robotic grasping and manipulation by enabling robots to interpret visual observations and natural language instru…
cs.RO2026
Beyond Visual Grasping: Benchmarking Complex Grasping from Detection to Execution
Hanyi Zhang, Khang Nguyen, Charith Munasinghe +10
Robust robotic grasping remains a fundamental challenge for complex real-world applications. Recent advances in large-scale models demonstrate promising capabilities for reasoning…