2 papers
cs.LG2023
Distance-rank Aware Sequential Reward Learning for Inverse Reinforcement Learning with Sub-optimal Demonstrations
Lu Li, Yuxin Pan, Ruobing Chen +4
Inverse reinforcement learning (IRL) aims to explicitly infer an underlying reward function based on collected expert demonstrations. Considering that obtaining expert demonstratio…
cs.LG2022
Backward Imitation and Forward Reinforcement Learning via Bi-directional Model Rollouts
Yuxin Pan, Fangzhen Lin
Traditional model-based reinforcement learning (RL) methods generate forward rollout traces using the learnt dynamics model to reduce interactions with the real environment. The re…