Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Modular Diffusion Policy Training: Decoupling and Recombining Guidance and Diffusion for Offline RL
Zhaoyang Chen, Cody Fleming
Classifier free guidance has shown strong potential in diffusion-based reinforcement learning. However, existing methods rely on joint training of the guidance module and the diffu…
cs.LG2024
FAWAC: Feasibility Informed Advantage Weighted Regression for Persistent Safety in Offline Reinforcement Learning
Prajwal Koirala, Zhanhong Jiang, Soumik Sarkar +1
Safe offline reinforcement learning aims to learn policies that maximize cumulative rewards while adhering to safety constraints, using only offline data for training. A key challe…
cs.LG2024
Towards Robust Car Following Dynamics Modeling via Blackbox Models: Methodology, Analysis, and Recommendations
Muhammad Bilal Shahid, Cody Fleming
The selection of the target variable is important while learning parameters of the classical car following models like GIPPS, IDM, etc. There is a vast body of literature on which…