2 papers
cs.LG2026
Learning the Model While Learning Q: Finite-Time Sample Complexity of Online SyncMBQ
Han-Dong Lim, HyeAnn Lee, Donghwan Lee
Reinforcement learning has witnessed significant advancements, particularly with the emergence of model-based approaches. Among these, -learning has proven to be a powerful algo…
cs.LG2024
Suppressing Overestimation in Q-Learning through Adversarial Behaviors
HyeAnn Lee, Donghwan Lee
The goal of this paper is to propose a new Q-learning algorithm with a dummy adversarial player, which is called dummy adversarial Q-learning (DAQ), that can effectively regulate t…