2 papers
cs.RO2026
Rethinking the Practicality of Vision-language-action Model: A Comprehensive Benchmark and An Improved Baseline
Wenxuan Song, Jiayi Chen, Xiaoquan Sun +12
Vision-Language-Action (VLA) models have emerged as a generalist robotic agent. However, existing VLAs are hindered by excessive parameter scales, prohibitive pre-training requirem…
cs.RO2025
Learning Generalizable Language-Conditioned Cloth Manipulation from Long Demonstrations
Hanyi Zhao, Jinxuan Zhu, Zihao Yan +3
Multi-step cloth manipulation is a challenging problem for robots due to the high-dimensional state spaces and the dynamics of cloth. Despite recent significant advances in end-to-…