1 paper · 1 filter
Shree Murthy, Rohan Pandey
We study two reproducible failure modes of deep multi-agent reinforcement learning in continuous-time pricing markets: (i) tacit cartel formation between competing DDPG agents, and…