7 papers
Scaling by Diversified Experience for Vision-Language-Action Models
Leiyu Wang, Zhaofengnian Wang, Xueqi Li +3
Vision-Language-Action models face significant challenges in real-world deployment due to the entanglement of high-level reasoning with low-level control, and the instability of po…
OpenEAI-Platform: An Open-source Embodied Artificial Intelligence Hardware-Software Unified Platform
Jinyuan Zhang, Luoyi Fan, Leiyu Wang +4
Embodied AI in the real world requires both accurate hardware and robust vision-language-action (VLA) policies. We present OpenEAI-Platform, a fully open-source platform that integ…
Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning
Jinzong Dong, Wei Huang, Jianshu Zhang +5
Offline reinforcement learning (RL), which optimizes policies using a previously collected static dataset, is an important branch of RL. A popular and promising approach is to regu…
V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation
Yaru Liu, Ao-bo Wang, Nanyang Ye
Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing scene generation methods often…
An End-to-end Flight Control Network for High-speed UAV Obstacle Avoidance based on Event-Depth Fusion
Dikai Shang, Jingyue Zhao, Shi Xu +2
Achieving safe, high-speed autonomous flight in complex environments with static, dynamic, or mixed obstacles remains challenging, as a single perception modality is incomplete. De…
V-CAGE: Context-Aware Generation and Verification for Scalable Long-Horizon Embodied Tasks
Yaru Liu, Ao-bo Wang, Nanyang Ye
Learning long-horizon embodied behaviors from synthetic data remains challenging because generated scenes are often physically implausible, language-driven programs frequently "suc…