activity
20242026
collaborators

11 papers

cs.RO2026

NativeMEM: Native Memory Compression for Long-Horizon Robotic Manipulation

Ziye Wang, Modi Shi, Chaojun Ni +5

How can pretrained Vision-Language-Action (VLA) models retain long-horizon visual histories with high-frequency updates without sacrificing efficiency? Existing approaches rely on…

cs.RO2026

RISE: Self-Improving Robot Policy with Compositional World Model

Jiazhi Yang, Kunyang Lin, Jinwei Li +10

Despite the sustained scaling on model capacity and data acquisition, Vision-Language-Action (VLA) models remain brittle in contact-rich and dynamic manipulation tasks, where minor…

cs.RO2026

HoloBrain-0 Technical Report

Xuewu Lin, Tianwei Lin, Yun Du +12

In this work, we introduce HoloBrain-0, a comprehensive Vision-Language-Action (VLA) framework that bridges the gap between foundation model research and reliable real-world robot…

cs.RO2025

SEM: Enhancing Spatial Understanding for Robust Robot Manipulation

Xuewu Lin, Tianwei Lin, Lichao Huang +4

A key challenge in robot manipulation lies in developing policy models with strong spatial understanding, the ability to reason about 3D geometry, object relations, and robot embod…

cs.RO2025

H-RDT: Human Manipulation Enhanced Bimanual Robotic Manipulation

Hongzhe Bi, Lingxuan Wu, Tianwei Lin +4

Imitation learning for robotic manipulation faces a fundamental challenge: the scarcity of large-scale, high-quality robot demonstration data. Recent robotic foundation models ofte…

cs.RO2025

FineGrasp: Towards Robust Grasping for Delicate Objects

Yun Du, Mengao Zhao, Tianwei Lin +3

Recent advancements in robotic grasping have led to its integration as a core module in many manipulation systems. For instance, language-driven semantic segmentation enables the g…