5 citations · 5 across the 2 of their papers we have counts for
1 paper · 1 filter
Aditya Kapoor, Sushant Swamy, Kale-ab Tessera +4
In multi-agent environments, agents often struggle to learn optimal policies due to sparse or delayed global rewards, particularly in long-horizon tasks where it is challenging to…