3 papers
cs.CR2026
VLALeaks: Membership Inference Attacks against Vision-Language-Action Models
Xukun Luan, Jinyan Liu, Xuesong Li +4
Vision-Language-Action (VLA) models enable end-to-end robot control and have garnered widespread attention. However, the memorization of training data inherent to VLA, coupled with…
cs.RO2026
GeoHAT: Geometry-Adaptive Hybrid Action Transformer for Mobile Manipulation
Xiangyu Zhu, Renjun Wu, Luzhou Ge +2
Whole-body mobile manipulation requires coordinating mobile base and manipulator under shifting viewpoints, posing challenges in geometric perception and action generation. Current…
cs.RO2026
ReMAP-DP: Reprojected Multi-view Aligned PointMaps for Diffusion Policy
Xinzhang Yang, Renjun Wu, Jinyan Liu +1
Generalist robot policies built upon 2D visual representations excel at semantic reasoning but inherently lack the explicit 3D spatial awareness required for high-precision tasks.…