1 paper
Liang Mi, Weijun Wang, Bowen Gao +12
Embodied reinforcement learning (RL) improves model capabilities with a pipeline of environment simulation, action generation, and model updates. These stages show heterogeneous CP…