1 paper
Ziqi Zhang, Jingzehua Xu, Zifeng Zhuang +4
Proximal Policy Optimization (PPO) has been broadly applied to robotics learning, showcasing stable training performance. However, the fixed clipping bound setting may limit the pe…