works on

From the 1 of 10 linked papers with an AI index.

activity
20242026
collaborators

10 papers

cs.RO2026

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation

Yao He, Gan Sun, Wenqi Liang +2

The paper introduces LifelongVLA, a framework that enables robots to continuously learn new manipulation tasks by using a dual-timescale adaptation mechanism and a cache-efficient…

cs.CV2026

Crafting Your Evolving Dreams: Concept-Incremental Versatile Customization

Jiahua Dong, Wenqi Liang, Hongliu Li +7

Custom diffusion models (CDMs) have garnered significant interest owing to their remarkable capacity for generating personalized concepts. However, the majority of CDMs unrealistic…

cs.RO2026

FocalPolicy: Frequency-Optimized Chunking and Locally Anchored Flow Matching for Coherent Visuomotor Policy

Qian He, Zhenshuo Yang, Wenqi Liang +3

Visuomotor policies aim to learn complex manipulation tasks from expert demonstrations. However, generating smooth and coherent trajectories remains challenging, as it requires bal…

cs.CV2026

PixelVLA: Advancing Pixel-level Understanding in Vision-Language-Action Model

Wenqi Liang, Gan Sun, Yao He +5

Vision-Language-Action models (VLAs) are emerging as powerful tools for learning generalizable visuomotor control policies. However, current VLAs are mostly trained on large-scale…

cs.RO2025

Never-Ending Behavior-Cloning Agent for Robotic Manipulation

Wenqi Liang, Gan Sun, Yao He +3

Relying on multi-modal observations, embodied robots (e.g., humanoid robots) could perform multiple robotic manipulation tasks in unstructured real-world environments. However, mos…

cs.CV2025

Bring Your Dreams to Life: Continual Text-to-Video Customization

Jiahua Dong, Xudong Wang, Wenqi Liang +7

Customized text-to-video generation (CTVG) has recently witnessed great progress in generating tailored videos from user-specific text. However, most CTVG methods assume that perso…