collaborators

5 papers

math.DS2025

Some remarks on stochastic converse Lyapunov theorems

Pavel Osinenko, Grigory Yaremenko

In this brief note, we investigate some constructions of Lyapunov functions for stochastic discrete-time stabilizable dynamical systems, in other words, controlled Markov chains. T…

cs.DC2025

SwarmRaft: Leveraging Consensus for Robust Drone Swarm Coordination in GNSS-Degraded Environments

Kapel Dev, Yash Madhwal, Sofia Shevelo +2

Unmanned aerial vehicle (UAV) swarms are increasingly used in critical applications such as aerial mapping, environmental monitoring, and autonomous delivery. However, the reliabil…

cs.LG2025

A universal policy wrapper with guarantees

Anton Bolychev, Georgiy Malaniya, Grigory Yaremenko +2

We introduce a universal policy wrapper for reinforcement learning agents that ensures formal goal-reaching guarantees. In contrast to standard reinforcement learning algorithms th…

cs.LG2025

Multi-CALF: A Policy Combination Approach with Statistical Guarantees

Georgiy Malaniya, Anton Bolychev, Grigory Yaremenko +2

We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL…

cs.LG2024

Comprehensive Overview of Reward Engineering and Shaping in Advancing Reinforcement Learning Applications

Sinan Ibrahim, Mostafa Mostafa, Ali Jnadi +2

The aim of Reinforcement Learning (RL) in real-world applications is to create systems capable of making autonomous decisions by learning from their environment through trial and e…