16 citations · 25 across the 18 of their papers we have counts for
1 paper · 1 filter
Ang Li, Xin Xu, Bin Liang +4
Reinforcement learning for user-centric agents is limited by the cost, latency, and risk of collecting online feedback, as well as by the lack of counterfactual comparisons under t…