collaborators

7 papers

cs.CV2026

Scaling by Diversified Experience for Vision-Language-Action Models

Leiyu Wang, Zhaofengnian Wang, Xueqi Li +3

Vision-Language-Action models face significant challenges in real-world deployment due to the entanglement of high-level reasoning with low-level control, and the instability of po…

cs.RO2026

OpenEAI-Platform: An Open-source Embodied Artificial Intelligence Hardware-Software Unified Platform

Jinyuan Zhang, Luoyi Fan, Leiyu Wang +4

Embodied AI in the real world requires both accurate hardware and robust vision-language-action (VLA) policies. We present OpenEAI-Platform, a fully open-source platform that integ…

cs.LG2026

Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning

Jinzong Dong, Wei Huang, Jianshu Zhang +5

Offline reinforcement learning (RL), which optimizes policies using a previously collected static dataset, is an important branch of RL. A popular and promising approach is to regu…

cs.RO2026

V-CAGE: Vision-Closed-Loop Agentic Generation Engine for Robotic Manipulation

Yaru Liu, Ao-bo Wang, Nanyang Ye

Scaling Vision-Language-Action (VLA) models requires massive datasets that are both semantically coherent and physically feasible. However, existing scene generation methods often…

cs.RO2026

An End-to-end Flight Control Network for High-speed UAV Obstacle Avoidance based on Event-Depth Fusion

Dikai Shang, Jingyue Zhao, Shi Xu +2

Achieving safe, high-speed autonomous flight in complex environments with static, dynamic, or mixed obstacles remains challenging, as a single perception modality is incomplete. De…

cs.RO2026

V-CAGE: Context-Aware Generation and Verification for Scalable Long-Horizon Embodied Tasks

Yaru Liu, Ao-bo Wang, Nanyang Ye

Learning long-horizon embodied behaviors from synthetic data remains challenging because generated scenes are often physically implausible, language-driven programs frequently "suc…