1 paper
Tung M. Luu, Donghoon Lee, Chang D. Yoo
Recent work in offline reinforcement learning (RL) has demonstrated the effectiveness of formulating decision-making as return-conditioned supervised learning. Notably, the decisio…