1 citations · 1 across the 2 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Structural Equivalence and Learning Dynamics in Delayed MARL
Jules Sintes, Ana Bušić, Jiamin Zhu
We formally establish the equivalence between Observation Delay (OD) and Action Delay (AD) in cooperative partially observable multi-agent systems using observation-action historie…
cs.LG2025★ 1 cited
WFCRL: A Multi-Agent Reinforcement Learning Benchmark for Wind Farm Control
Claire Bizon Monroc, Ana Bušić, Donatien Dubuc +1
The wind farm control problem is challenging, since conventional model-based control strategies require tractable models of complex aerodynamical interactions between the turbines…
cs.LG2024
Reinforcement Learning and Regret Bounds for Admission Control
Lucas Weber, Ana Bušić, Jiamin Zhu
The expected regret of any reinforcement learning algorithm is lower bounded by for undiscounted returns, where is the diameter of the Markov decisi…