7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.LG2024★ 7 cited
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
Chenjia Bai, Lingxiao Wang, Jianye Hao +4
Offline Reinforcement Learning (RL) has shown promising results in learning a task-specific policy from a fixed dataset. However, successful offline RL often relies heavily on the…
stat.ML2023
Exploration in Model-based Reinforcement Learning with Randomized Reward
Lingxiao Wang, Ping Li
Model-based Reinforcement Learning (MBRL) has been widely adapted due to its sample efficiency. However, existing worst-case regret analysis typically requires optimistic planning,…