Showing cs.ROShow all
3 papers · 1 filter
cs.RO2026
ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning
Tuan Van Vo, Tan Q. Nguyen, Khang Nguyen +5
Vision-Language-Action (VLA) models have gained much attention from the research community thanks to their strength in translating multimodal observations with linguistic instructi…
cs.RO2025
ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning
Tuan Van Vo, Tan Quang Nguyen, Khang Minh Nguyen +2
Vision-Language-Action (VLA) models have gained much attention from the research community thanks to their strength in translating multimodal observations with linguistic instructi…
cs.RO2024
Language-driven Grasp Detection with Mask-guided Attention
Tuan Van Vo, Minh Nhat Vu, Baoru Huang +4
Grasp detection is an essential task in robotics with various industrial applications. However, traditional methods often struggle with occlusions and do not utilize language for g…