4 papers
ReinVBC: A Model-based Reinforcement Learning Approach to Vehicle Braking Controller
Haoxin Lin, Junjie Zhou, Daheng Xu +1
Braking system, the key module to ensure the safety and steer-ability of current vehicles, relies on extensive manual calibration during production. Reducing labor and time consump…
Off-Policy Value-Based Reinforcement Learning for Large Language Models
Peng-Yuan Wang, Ziniu Li, Tian Xu +8
Improving data utilization efficiency is critical for scaling reinforcement learning (RL) for long-horizon tasks where generating trajectories is expensive. However, the dominant R…
VLGOR: Visual-Language Knowledge Guided Offline Reinforcement Learning for Generalizable Agents
Pengsen Liu, Maosen Zeng, Nan Tang +4
Combining Large Language Models (LLMs) with Reinforcement Learning (RL) enables agents to interpret language instructions more effectively for task execution. However, LLMs typical…
Generalist Reward Models: Found Inside Large Language Models
Yi-Chen Li, Tian Xu, Yang Yu +6
The alignment of Large Language Models (LLMs) is critically dependent on reward models trained on costly human preference data. While recent work explores bypassing this cost with…