2 papers
cs.RO2025
LLaDA-VLA: Vision Language Diffusion Action Models
Yuqing Wen, Hebei Li, Kefan Gu +3
The rapid progress of auto-regressive vision-language models (VLMs) has inspired growing interest in vision-language-action models (VLA) for robotic manipulation. Recently, masked…
cs.RO2025
ROSA: Harnessing Robot States for Vision-Language and Action Alignment
Yuqing Wen, Kefan Gu, Haoxuan Liu +4
Vision-Language-Action (VLA) models have recently made significant advance in multi-task, end-to-end robotic control, due to the strong generalization capabilities of Vision-Langua…