2 citations · 5 across the 7 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
Offline Meta Reinforcement Learning with In-Distribution Online Adaptation
Jianhao Wang, Jin Zhang, Haozhe Jiang +3
Recent offline meta-reinforcement learning (meta-RL) methods typically utilize task-dependent behavior policies (e.g., training RL agents on each individual task) to collect a mult…
cs.LG2022★ 2 cited
A Near-Optimal Primal-Dual Method for Off-Policy Learning in CMDP
Fan Chen, Junyu Zhang, Zaiwen Wen
As an important framework for safe Reinforcement Learning, the Constrained Markov Decision Process (CMDP) has been extensively studied in the recent literature. However, despite th…