8 citations · 10 across the 3 of their papers we have counts for
3 papers
cs.LG2022★ 1 cited
Insights From the NeurIPS 2021 NetHack Challenge
Eric Hambro, Sharada Mohanty, Dmitrii Babaev +26
In this report, we summarize the takeaways from the first NeurIPS 2021 NetHack Challenge. Participants were tasked with developing a program or agent that can win (i.e., 'ascend' i…
cs.LG2021★ 8 cited
Automated Learning Rate Scheduler for Large-batch Training
Chiheon Kim, Saehoon Kim, Jongmin Kim +2
Large-batch training has been essential in leveraging large-scale datasets and models in deep learning. While it is computationally beneficial to use large batch sizes, it often re…
cs.LG2020★ 1 cited
Entropy-Augmented Entropy-Regularized Reinforcement Learning and a Continuous Path from Policy Gradient to Q-Learning
Donghoon Lee
Entropy augmented to reward is known to soften the greedy argmax policy to softmax policy. Entropy augmentation is reformulated and leads to a motivation to introduce an additional…