Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
Disentangling Visuo-Tactile Foresight: Oracle-Guided Interface Discovery for World Action Models
Zihang Yao, Chaoyue Ding, Yingying Yu
Contact-rich manipulation remains challenging because successful control depends on physical interaction cues that are often weakly observable from vision alone. Recent tactile wor…
cs.RO2026
SAM3D-Guided Object-Centric Representation Alignment for Vision-Language-Action Models
Zonghe Liu, Shanyuan Jie, Xiaoquan Sun +4
Vision-Language-Action (VLA) models have shown strong potential for general robot manipulation, but most existing models rely on 2D visual-language backbones and lack fine-grained…