3 papers
cs.LG2026
EMAgnet: Parameter-Space EMA Regularization for Policy Gradient Self-Play in Large Games
Tristan Maidment, JB Lanier, Chase McDonald +5
Recent work has established that regularized policy gradient methods such as PPO, when used in self-play, can match or exceed specialized game-theoretic algorithms for solving two-…
cs.HC2026
CoGrid & the Multi-User Gymnasium: A Framework for Multi-Agent Experimentation
Chase McDonald, Cleotilde Gonzalez
The increasing integration of artificial intelligence (AI) in everyday life brings with it new challenges and questions for regarding how humans interact with autonomous agents. Mu…
cs.HC2025
Controllable Complementarity: Subjective Preferences in Human-AI Collaboration
Chase McDonald, Cleotilde Gonzalez
Research on human-AI collaboration often prioritizes objective performance. However, understanding human subjective preferences is essential to improving human-AI complementarity a…