1 paper · 1 filter
Ziping Xu, Kelly W. Zhang, Susan A. Murphy
Online Reinforcement Learning (RL) is typically framed as the process of minimizing cumulative regret (CR) through interactions with an unknown environment. However, real-world RL…