1 citations · 1 across the 6 of their papers we have counts for
Showing cs.ROShow all
3 papers · 1 filter
cs.RO2025
IntentionVLA: Generalizable and Efficient Embodied Intention Reasoning for Human-Robot Interaction
Yandu Chen, Kefan Gu, Yuqing Wen +3
Vision-Language-Action (VLA) models leverage pretrained vision-language models (VLMs) to couple perception with robotic control, offering a promising path toward general-purpose em…
cs.RO2025
LLaDA-VLA: Vision Language Diffusion Action Models
Yuqing Wen, Hebei Li, Kefan Gu +3
The rapid progress of auto-regressive vision-language models (VLMs) has inspired growing interest in vision-language-action models (VLA) for robotic manipulation. Recently, masked…
cs.RO2025
ROSA: Harnessing Robot States for Vision-Language and Action Alignment
Yuqing Wen, Kefan Gu, Haoxuan Liu +4
Vision-Language-Action (VLA) models have recently made significant advance in multi-task, end-to-end robotic control, due to the strong generalization capabilities of Vision-Langua…