2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 2 cited
Active Reward Learning from Multiple Teachers
Peter Barnett, Rachel Freedman, Justin Svegliato +1
Reward learning algorithms utilize human feedback to infer a reward function, which is then used to train an AI system. This human feedback is often a preference comparison, in whi…
cs.GT2022
For Learning in Symmetric Teams, Local Optima are Global Nash Equilibria
Scott Emmons, Caspar Oesterheld, Andrew Critch +2
Although it has been known since the 1970s that a globally optimal strategy profile in a common-payoff game is a Nash equilibrium, global optimality is a strict requirement that li…