1 paper
Jun Xue, Junze Wang, Shanze Wang +3
Scaling Maximum Entropy Reinforcement Learning (RL) to high-dimensional humanoid control remains a fundamental challenge, as the ''curse of dimensionality'' induces severe explorat…