Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Independent RL for Cooperative-Competitive Agents: A Mean-Field Perspective
Muhammad Aneeq uz Zaman, Alec Koppel, Mathieu Laurière +1
We address in this paper Reinforcement Learning (RL) among agents that are grouped into teams such that there is cooperation within each team but general-sum (non-zero sum) competi…
cs.LG2023
Byzantine-Resilient Decentralized Multi-Armed Bandits
Jingxuan Zhu, Alec Koppel, Alvaro Velasquez +1
In decentralized cooperative multi-armed bandits (MAB), each agent observes a distinct stream of rewards, and seeks to exchange information with others to select a sequence of arms…