1 paper
Suei-Wen Chen, Keith Ross, Pierre Youssef
Monte Carlo Exploring Starts (MCES), which aims to learn the optimal policy using only sample returns, is a simple and natural algorithm in reinforcement learning which has been sh…