5 papers
Some remarks on stochastic converse Lyapunov theorems
Pavel Osinenko, Grigory Yaremenko
In this brief note, we investigate some constructions of Lyapunov functions for stochastic discrete-time stabilizable dynamical systems, in other words, controlled Markov chains. T…
SwarmRaft: Leveraging Consensus for Robust Drone Swarm Coordination in GNSS-Degraded Environments
Kapel Dev, Yash Madhwal, Sofia Shevelo +2
Unmanned aerial vehicle (UAV) swarms are increasingly used in critical applications such as aerial mapping, environmental monitoring, and autonomous delivery. However, the reliabil…
A universal policy wrapper with guarantees
Anton Bolychev, Georgiy Malaniya, Grigory Yaremenko +2
We introduce a universal policy wrapper for reinforcement learning agents that ensures formal goal-reaching guarantees. In contrast to standard reinforcement learning algorithms th…
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
Georgiy Malaniya, Anton Bolychev, Grigory Yaremenko +2
We introduce Multi-CALF, an algorithm that intelligently combines reinforcement learning policies based on their relative value improvements. Our approach integrates a standard RL…
Comprehensive Overview of Reward Engineering and Shaping in Advancing Reinforcement Learning Applications
Sinan Ibrahim, Mostafa Mostafa, Ali Jnadi +2
The aim of Reinforcement Learning (RL) in real-world applications is to create systems capable of making autonomous decisions by learning from their environment through trial and e…