1 citations · 1 across the 3 of their papers we have counts for
3 papers
math.OC2021
A study of first-passage time minimization via Q-learning in heated gridworlds
M. A. Larchenko, P. Osinenko, G. Yaremenko +1
Optimization of first-passage times is required in applications ranging from nanobots navigation to market trading. In such settings, one often encounters unevenly distributed nois…
cs.RO2021★ 1 cited
An experimental study of two predictive reinforcement learning methods and comparison with model-predictive control
Dmitrii Dobriborsci, Pavel Osinenko
Reinforcement learning (RL) has been successfully used in various simulations and computer games. Industry-related applications, such as autonomous mobile robot motion control, are…
math.DS2021
Effects of sampling and horizon in predictive reinforcement learning
Pavel Osinenko, Dmitrii Dobriborsci
Plain reinforcement learning (RL) may be prone to loss of convergence, constraint violation, unexpected performance, etc. Commonly, RL agents undergo extensive learning stages to a…