1 paper · 1 filter
Aaditya Mehta, Arya Shah
Cooperative multi-agent reinforcement learning often adds social terms to individual rewards, yet the scale of those terms is usually chosen by hand. We ask whether a guilt signal…