Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Structural Equivalence and Learning Dynamics in Delayed MARL
Jules Sintes, Ana BuÅ¡iÄ, Jiamin Zhu
We formally establish the equivalence between Observation Delay (OD) and Action Delay (AD) in cooperative partially observable multi-agent systems using observation-action historie…
cs.LG2025
WFCRL: A Multi-Agent Reinforcement Learning Benchmark for Wind Farm Control
Claire Bizon Monroc, Ana BuÅ¡iÄ, Donatien Dubuc +1
The wind farm control problem is challenging, since conventional model-based control strategies require tractable models of complex aerodynamical interactions between the turbines…
cs.LG2024
Reinforcement Learning and Regret Bounds for Admission Control
Lucas Weber, Ana BuÅ¡iÄ, Jiamin Zhu
The expected regret of any reinforcement learning algorithm is lower bounded by for undiscounted returns, where is the diameter of the Markov decis…