3 papers
cs.LG2025
A universal policy wrapper with guarantees
Anton Bolychev, Georgiy Malaniya, Grigory Yaremenko +2
We introduce a universal policy wrapper for reinforcement learning agents that ensures formal goal-reaching guarantees. In contrast to standard reinforcement learning algorithms th…
cs.LG2025
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
Georgiy Malaniya, Anton Bolychev, Grigory Yaremenko +2
We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL…
eess.SY2022
A framework for online, stabilizing reinforcement learning
Grigory Yaremenko, Georgiy Malaniya, Pavel Osinenko
Online reinforcement learning is concerned with training an agent on-the-fly via dynamic interaction with the environment. Here, due to the specifics of the application, it is not…