4 papers · 1 filter
WSA: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control
Jiahao Jiang, Jianing Zhang, Zhenhan Yin +8
Recent advances in embodied AI have established robot foundation models (RFMs) as the dominant approach for generalist robotic systems to date. By leveraging imitation learning on…
DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models
Siyuan Xu, Tianshi Wang, Fengling Li +2
Vision-Language-Action models (VLAs) have demonstrated strong potential for embodied AI, yet their deployment on resource-limited robots remains challenging due to high memory and…
Non-Markovian Long-Horizon Robot Manipulation via Keyframe Chaining
Yipeng Chen, Wentao Tan, Lei Zhu +4
Existing Vision-Language-Action (VLA) models often struggle to generalize to long-horizon tasks due to their heavy reliance on immediate observations. While recent studies incorpor…
Sim-and-Human Co-training for Data-Efficient and Generalizable Robotic Manipulation
Kaipeng Fang, Weiqing Liang, Yuyang Li +5
Synthetic simulation data and real-world human data provide scalable alternatives to circumvent the prohibitive costs of robot data collection. However, these sources suffer from t…