2 papers
cs.LG2024
vMFER: Von Mises-Fisher Experience Resampling Based on Uncertainty of Gradient Directions for Policy Improvement
Yiwen Zhu, Jinyi Liu, Wenya Wei +7
Reinforcement Learning (RL) is a widely employed technique in decision-making problems, encompassing two fundamental operations -- policy evaluation and policy improvement. Enhanci…
math.OC2021
A modified limited memory Nesterov's accelerated quasi-Newton
S. Indrapriyadarsini, Shahrzad Mahboubi, Hiroshi Ninomiya +2
The Nesterov's accelerated quasi-Newton (L)NAQ method has shown to accelerate the conventional (L)BFGS quasi-Newton method using the Nesterov's accelerated gradient in several neur…