collaborators

12 papers

cs.AI2026

UniMM: A Unified Mixture Model Framework for Multi-Agent Simulation

Longzhong Lin, Xuewu Lin, Kechun Xu +4

Simulation plays a crucial role in assessing autonomous driving systems, where the generation of realistic multi-agent behaviors is a key aspect. In multi-agent simulation, the pri…

cs.RO2026

BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots

Yufei Wei, Chenxiao Hu, Wangtao Lu +5

Scale-consistent ego-motion estimation is fundamental for autonomous ground robots. Bird's-Eye-View (BEV) representation naturally addresses the scale drift problem of monocular vi…

cs.RO2026

IntentReact: Guiding Reactive Object-Centric Navigation via Topological Intent

Yanmei Jiao, Anpeng Lu, Wenhan Hu +4

Object-goal visual navigation requires robots to reason over semantic structure and act effectively under partial observability. Recent approaches based on object-level topological…

cs.RO2026

Direction Matters: Learning Force Direction Enables Sim-to-Real Contact-Rich Manipulation

Yifei Yang, Anzhe Chen, Zhenjie Zhu +6

Sim-to-real transfer for contact-rich manipulation remains challenging due to the inherent discrepancy in contact dynamics. While existing methods often rely on costly real-world d…

cs.RO2025

ETP-R1: Evolving Topological Planning with Reinforcement Fine-tuning for Vision-Language Navigation in Continuous Environments

Shuhao Ye, Sitong Mao, Yuxiang Cui +6

Vision-Language Navigation in Continuous Environments (VLN-CE) requires an embodied agent to navigate towards target in continuous environments, following natural language instruct…

cs.RO2025

Seeing to Act, Prompting to Specify: A Bayesian Factorization of Vision Language Action Policy

Kechun Xu, Zhenjie Zhu, Anzhe Chen +7

The pursuit of out-of-distribution generalization in Vision-Language-Action (VLA) models is often hindered by catastrophic forgetting of the Vision-Language Model (VLM) backbone du…