4 citations · 4 across the 1 of their papers we have counts for
2 papers
cs.LG2020
Explainability for fair machine learning
Tom Begley, Tobias Schwedes, Christopher Frye +1
As the decisions made or influenced by machine learning models increasingly impact our lives, it is crucial to detect, understand, and mitigate unfairness. But even simply determin…
cs.AI2019★ 4 cited
Parenting: Safe Reinforcement Learning from Human Input
Christopher Frye, Ilya Feige
Autonomous agents trained via reinforcement learning present numerous safety concerns: reward hacking, negative side effects, and unsafe exploration, among others. In the context o…