works on

From the 1 of 7 linked papers with an AI index.

collaborators

7 papers

cs.RO2026

Explicit Kinematic Guidance from Analytic Concepts for Vision-Language-Action Models

Mingyang Sun, Jiude Wei, Xiujian Liang +4

The paper introduces a Concept Expert module that extracts 3D kinematic and structural information to create Analytic Concepts, providing explicit guidance for Vision-Language-Acti…

cs.RO2026

Analytic Concept-Centric Memory for Agentic Embodied Manipulation

Mingyang Sun, Xiujian Liang, Jiude Wei +4

Long-horizon embodied manipulation requires agents to remember persistent objects, track changing scene states, and reuse prior interaction knowledge. However, existing agent memor…

cs.RO2026

TrajBooster: Boosting Humanoid Whole-Body Manipulation via Trajectory-Centric Learning

Jiacheng Liu, Pengxiang Ding, Qihang Zhou +8

Recent Vision-Language-Action models show potential to generalize across embodiments but struggle to quickly align with a new robot's action space when high-quality demonstrations…

cs.RO2026

Rethinking the Practicality of Vision-language-action Model: A Comprehensive Benchmark and An Improved Baseline

Wenxuan Song, Jiayi Chen, Xiaoquan Sun +12

Vision-Language-Action (VLA) models have emerged as a generalist robotic agent. However, existing VLAs are hindered by excessive parameter scales, prohibitive pre-training requirem…

cs.RO2025

L1 Sample Flow for Efficient Visuomotor Learning

Weixi Song, Zhetao Chen, Tao Xu +6

Denoising-based models, such as diffusion and flow matching, have been a critical component of robotic manipulation for their strong distribution-fitting and scaling capacity. Conc…

cs.LG2025

Iterative Refinement of Flow Policies in Probability Space for Online Reinforcement Learning

Mingyang Sun, Pengxiang Ding, Weinan Zhang +1

While behavior cloning with flow/diffusion policies excels at learning complex skills from demonstrations, it remains vulnerable to distributional shift, and standard RL methods st…