1 paper
Denis Belomestny, Alexander Gasnikov, Egor Gladin +5
Reinforcement learning (RL) is increasingly grounded in tools from probability, optimization, and operator theory. This survey organizes the mathematical structures that underpin t…