1 paper · 1 filter
Yujie Shen, Haowen Chen
This paper introduces PPO-HSC (Proximal Policy Optimization with High-order Sampling Coverage), an exploratory reinforcement learning framework designed to address the "Invisible S…