From the 1 of 2 linked papers with an AI index.
3 papers · 1 filter
TADPO: Reinforcement Learning Goes Off-road
Zhouchonghao Wu, Raymond Song, Vedant Mundheda +3
The paper introduces TADPO, a policy‑gradient method that extends PPO with teacher‑student guidance, and uses it in a vision‑based end‑to‑end reinforcement learning system for high…
Planning with Adaptive World Models for Autonomous Driving
Arun Balajee Vasudevan, Neehar Peri, Jeff Schneider +1
Motion planning is crucial for safe navigation in complex urban environments. Historically, motion planners (MPs) have been evaluated with procedurally-generated simulators like CA…
Tractable Joint Prediction and Planning over Discrete Behavior Modes for Urban Driving
Adam Villaflor, Brian Yang, Huangyuan Su +3
Significant progress has been made in training multimodal trajectory forecasting models for autonomous driving. However, effectively integrating these models with downstream planne…