3 papers
cs.LG2026
An Agency-Transferring Model-Free Policy Enhancement Technique
Anton Bolychev, Georgiy Malaniya, Sinan Ibrahim +1
Training reinforcement learning (RL) policies from scratch is costly: it requires careful reward and environment design, extensive tuning, and substantial computation. Yet many con…
cs.LG2025
A universal policy wrapper with guarantees
Anton Bolychev, Georgiy Malaniya, Grigory Yaremenko +2
We introduce a universal policy wrapper for reinforcement learning agents that ensures formal goal-reaching guarantees. In contrast to standard reinforcement learning algorithms th…
cs.LG2025
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
Georgiy Malaniya, Anton Bolychev, Grigory Yaremenko +2
We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL…