6 papers
Driving Like Yourself: A Benchmark for Closed-Loop Personalized End-to-End Autonomous Driving
Xiaoru Dong, Ruiqin Li, Xiao Han +7
Human driving behavior is inherently diverse, yet most end-to-end autonomous driving (E2E-AD) systems learn a single average driving style, neglecting individual differences. Achie…
Potential-Guided Flow Matching for Vision-Language-Action Policy Improvement
Yunpeng Mei, Jiakai He, Hongjie Cao +12
Large vision-language-action (VLA) policies are increasingly trained as conditional generative models over action chunks. Yet deployment produces mixed-quality experience-successfu…
Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model
Hanqing Wang, Shaoyang Wang, Yiming Zhong +7
Affordance grounding focuses on predicting the specific regions of objects that are associated with the actions to be performed by robots. It plays a vital role in the fields of hu…
UniBioTransfer: A Unified Framework for Multiple Biometrics Transfer
Caiyi Sun, Yujing Sun, Xiangyu Li +5
Deepface generation has traditionally followed a task-driven paradigm, where distinct tasks (e.g., face transfer and hair transfer) are addressed by task-specific models. Neverthel…
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation
Jiamin Wang, Yichen Yao, Xiang Feng +5
The generation of temporally consistent, high-fidelity driving videos over extended horizons presents a fundamental challenge in autonomous driving world modeling. Existing approac…
RealDex: Towards Human-like Grasping for Robotic Dexterous Hand
Yumeng Liu, Yaxun Yang, Youzhuo Wang +9
In this paper, we introduce RealDex, a pioneering dataset capturing authentic dexterous hand grasping motions infused with human behavioral patterns, enriched by multi-view and mul…