3 papers
cs.RO2025
VTLA: Vision-Tactile-Language-Action Model with Preference Learning for Insertion Manipulation
Chaofan Zhang, Peng Hao, Xiaoge Cao +3
While vision-language models have advanced significantly, their application in language-conditioned robotic manipulation is still underexplored, especially for contact-rich tasks t…
cs.RO2025
CLTP: Contrastive Language-Tactile Pre-training for 3D Contact Geometry Understanding
Wenxuan Ma, Xiaoge Cao, Yixiang Zhang +7
Recent advancements in integrating tactile sensing with vision-language models (VLMs) have demonstrated remarkable potential for robotic multimodal perception. However, existing ta…
cs.RO2025
TLA: Tactile-Language-Action Model for Contact-Rich Manipulation
Peng Hao, Chaofan Zhang, Dingzhe Li +4
Significant progress has been made in vision-language models. However, language-conditioned robotic manipulation for contact-rich tasks remains underexplored, particularly in terms…