activity
20242026
collaborators

5 papers

cs.CL2026

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

Qiyao Ma, Dechen Gao, Rui Cai +4

Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a central mechanism for capturing d…

cs.CV2025

VITA: Vision-to-Action Flow Matching Policy

Dechen Gao, Boqi Zhao, Andrew Lee +6

Conventional flow matching and diffusion-based policies sample via iterative denoising from standard noise distributions (e.g., Gaussian), and require conditioning modules to repea…

cs.RO2025

IN-RIL: Interleaved Reinforcement and Imitation Learning for Policy Fine-Tuning

Dechen Gao, Hang Wang, Hanchu Zhou +5

Imitation learning (IL) and reinforcement learning (RL) each offer distinct advantages for robotics policy learning: IL provides stable learning from demonstrations, and RL promote…

cs.RO2024

EI-Drive: A Platform for Cooperative Perception with Realistic Communication Models

Hanchu Zhou, Edward Xie, Wei Shao +3

The growing interest in autonomous driving calls for realistic simulation platforms capable of accurately simulating cooperative perception process in realistic traffic scenarios.…

cs.RO2024

CarDreamer: Open-Source Learning Platform for World Model based Autonomous Driving

Dechen Gao, Shuangyu Cai, Hanchu Zhou +3

To safely navigate intricate real-world scenarios, autonomous vehicles must be able to adapt to diverse road conditions and anticipate future events. World model (WM) based reinfor…