1 paper
Johan Obando-Ceron, Lu Li, Scott Fujimoto +3
Scaling reinforcement learning (RL) to diverse multitask settings remains a central challenge. While recent advances in model-based RL achieve strong performance, they rely on plan…