Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Streaming Reinforcement Learning under Partial Observability with Real-Time Recurrent Learning
Noah Farr, Aryaman Reddi, Carlo D'Eramo +1
Streaming reinforcement learning has emerged as an online learning paradigm that conforms to the restrictions of natural learning agents that process data incrementally, i.e. with…
cs.LG2025
Deep Learning Agents Trained For Avoidance Behave Like Hawks And Doves
Aryaman Reddi
We present heuristically optimal strategies expressed by deep learning agents playing a simple avoidance game. We analyse the learning and behaviour of two agents within a symmetri…
cs.LG2025
-Level Policy Gradients for Multi-Agent Reinforcement Learning
Aryaman Reddi, Gabriele Tiboni, Jan Peters +1
Actor-critic algorithms for deep multi-agent reinforcement learning (MARL) typically employ a policy update that responds to the current strategies of other agents. While being str…