3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 3 cited
Efficient Methods for Non-stationary Online Learning
Peng Zhao, Yan-Feng Xie, Lijun Zhang +1
Non-stationary online learning has drawn much attention in recent years. In particular, dynamic regret and adaptive regret are proposed as two principled performance measures for o…
cs.LG2022
Dynamic Regret of Online Markov Decision Processes
Peng Zhao, Long-Fei Li, Zhi-Hua Zhou
We investigate online Markov Decision Processes (MDPs) with adversarially changing loss functions and known transitions. We choose dynamic regret as the performance measure, define…