Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems
Wesley A. Suttle, Vipul K. Sharma, Krishna C. Kosaraju +4
We develop provably safe and convergent reinforcement learning (RL) algorithms for control of nonlinear dynamical systems, bridging the gap between the hard safety guarantees of co…
cs.LG2023
Byzantine-Resilient Decentralized Multi-Armed Bandits
Jingxuan Zhu, Alec Koppel, Alvaro Velasquez +1
In decentralized cooperative multi-armed bandits (MAB), each agent observes a distinct stream of rewards, and seeks to exchange information with others to select a sequence of arms…