2 papers
cs.LG2026
EMAgnet: Parameter-Space EMA Regularization for Policy Gradient Self-Play in Large Games
Tristan Maidment, JB Lanier, Chase McDonald +5
Recent work has established that regularized policy gradient methods such as PPO, when used in self-play, can match or exceed specialized game-theoretic algorithms for solving two-…
cs.HC2024
Human-like Bots for Tactical Shooters Using Compute-Efficient Sensors
Niels Justesen, Maria Kaselimi, Sam Snodgrass +12
Artificial intelligence (AI) has enabled agents to master complex video games, from first-person shooters like Counter-Strike to real-time strategy games such as StarCraft II and r…