1 citations · 1 across the 1 of their papers we have counts for
1 paper
Kihyuk Hong, Yuhang Li, Ambuj Tewari
Offline constrained reinforcement learning (RL) aims to learn a policy that maximizes the expected cumulative reward subject to constraints on expected cumulative cost using an exi…