3 citations · 6 across the 3 of their papers we have counts for
Showing 2020Show all
2 papers · 1 filter
cs.LG2020
Cooperative-Competitive Reinforcement Learning with History-Dependent Rewards
Keyang He, Bikramjit Banerjee, Prashant Doshi
Consider a typical organization whose worker agents seek to collectively cooperate for its general betterment. However, each individual agent simultaneously seeks to act to secure…
cs.LG2020★ 2 cited
Maximum Entropy Multi-Task Inverse RL
Saurabh Arora, Bikramjit Banerjee, Prashant Doshi
Multi-task IRL allows for the possibility that the expert could be switching between multiple ways of solving the same problem, or interleaving demonstrations of multiple tasks. Th…