3 citations · 3 across the 3 of their papers we have counts for
1 paper · 1 filter
Dhruv Malik, Yuanzhi Li, Aarti Singh
Policy regret is a well established notion of measuring the performance of an online learning algorithm against an adaptive adversary. We study restrictions on the adversary that e…