8 citations · 17 across the 7 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2023
Diverse Policies Converge in Reward-free Markov Decision Processe
Fanqi Lin, Shiyu Huang, Weiwei Tu
Reinforcement learning has achieved great success in many decision-making tasks, and traditional reinforcement learning algorithms are mainly designed for obtaining a single optima…
cs.LG2019★ 2 cited
SAdam: A Variant of Adam for Strongly Convex Functions
Guanghui Wang, Shiyin Lu, Weiwei Tu +1
The Adam algorithm has become extremely popular for large-scale machine learning. Under convexity condition, it has been proved to enjoy a data-dependant regret bound…