Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
UVIP: Model-Free Approach to Evaluate Reinforcement Learning Algorithms
Denis Belomestny, Ilya Levin, Alexey Naumov +1
Policy evaluation is an important instrument for the comparison of different algorithms in Reinforcement Learning (RL). However, even a precise knowledge of the value function $V^Ï…
cs.LG2024
Improving GFlowNets with Monte Carlo Tree Search
Nikita Morozov, Daniil Tiapkin, Sergey Samsonov +2
Generative Flow Networks (GFlowNets) treat sampling from distributions over compositional discrete spaces as a sequential decision-making problem, training a stochastic policy to c…