1 paper
Yuanfu Wang, Chao Yang, Ying Wen +2
Recent advancements in offline reinforcement learning (RL) have underscored the capabilities of Return-Conditioned Supervised Learning (RCSL), a paradigm that learns the action dis…