3 papers
cs.LG2026
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
Cheng Yin, Yankai Lin, Wang Xu +4
Does Chain-of-Thought (CoT) reasoning genuinely improve Vision Language Action (VLA) models, or does it merely add overhead? Existing CoT-VLA systems report limited and inconsisten…
cs.RO2026
Humanizing Robot Gaze Shifts: A Framework for Natural Gaze Shifts in Humanoid Robots
Jingchao Wei, Jingkai Qin, Yuxiao Cao +4
Leveraging auditory and visual feedback for attention reorientation is essential for natural gaze shifts in social interaction. However, enabling humanoid robots to perform natural…
cs.RO2025
Whole-body Motion Control of an Omnidirectional Wheel-Legged Mobile Manipulator via Contact-Aware Dynamic Optimization
Zong Chen, Shaoyang Li, Ben Liu +3
Wheel-legged robots with integrated manipulators hold great promise for mobile manipulation in logistics, industrial automation, and human-robot collaboration. However, unified con…