1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.AI2025
Enhancing Multi-Agent Collaboration with Attention-Based Actor-Critic Policies
Hugo Garrido-Lestache Belinchon, Jeremy Kedziora
This paper introduces Team-Attention-Actor-Critic (TAAC), a reinforcement learning algorithm designed to enhance multi-agent collaboration in cooperative environments. TAAC employs…
cs.AI2025
Strategy Masking: A Method for Guardrails in Value-based Reinforcement Learning Agents
Jonathan Keane, Sam Keyser, Jeremy Kedziora
The use of reward functions to structure AI learning and decision making is core to the current reinforcement learning paradigm; however, without careful design of reward functions…
cs.LG2024★ 1 cited
Prediction Instability in Machine Learning Ensembles
Jeremy Kedziora
In machine learning ensembles predictions from multiple models are aggregated. Despite widespread use and strong performance of ensembles in applied problems little is known about…