12 citations · 15 across the 3 of their papers we have counts for
3 papers
Challenges for Using Impact Regularizers to Avoid Negative Side Effects
David Lindner, Kyle Matoba, Alexander Meulemans
Designing reward functions for reinforcement learning is difficult: besides specifying which behavior is rewarded for a task, the reward also has to discourage undesired outcomes.…
Detecting Spiky Corruption in Markov Decision Processes
Jason Mancuso, Tomasz Kisielewski, David Lindner +1
Current reinforcement learning methods fail if the reward function is imperfect, i.e. if the agent observes reward different from what it actually receives. We study this problem w…
Sensing Social Media Signals for Cryptocurrency News
Johannes Beck, Roberta Huang, David Lindner +4
The ability to track and monitor relevant and important news in real-time is of crucial interest in multiple industrial sectors. In this work, we focus on the set of cryptocurrency…