From the 1 of 4 linked papers with an AI index.
4 papers
On the Identifiability of Controlled World Models
Xiangteng Zhang, Yang Guan, Bo Zhang +3
World model serves as a promising tool to infer environment dynamics under high-dimensional observations and candidate actions. Recently, LeCun's JEPA provides a compelling framewo…
FAST: A Framework for Aligned Sampling and Training in Parallel Reinforcement Learning for Autonomous Driving
Bonan Wang, Letian Tao, Bin Shuai +7
The paper introduces FAST, a synchronous parallel framework that improves sampling efficiency for deep reinforcement learning in autonomous driving by aligning parallel simulations…
STAPO: Stabilizing Reinforcement Learning for LLMs by Silencing Rare Spurious Tokens
Shiqi Liu, Zeyu He, Guojian Zhan +10
Reinforcement Learning (RL) has significantly improved large language model reasoning, but existing RL fine-tuning methods rely heavily on heuristic techniques such as entropy regu…
Towards Physically Consistent 4D Scene Reconstruction for Closed-loop Autonomous Driving Simulation
Bowyn Tan, Yutong Xie, Bai Huang +5
High-fidelity street scene reconstruction is pivotal for end-to-end autonomous driving simulation, where novel-view synthesis (NVS) and time-varying information modeling are two fu…