2 papers
cs.LG2026
Performance Variation in Deep Reinforcement Learning
Haruto Tanaka, A. Rupam Mahmood
Deep reinforcement learning (RL) algorithms often suffer from low run-to-run robustness, manifesting as significant performance variation across independent runs of identically con…
cs.LG2024
Directions of Curvature as an Explanation for Loss of Plasticity
Alex Lewandowski, Haruto Tanaka, Dale Schuurmans +1
Loss of plasticity is a phenomenon in which neural networks lose their ability to learn from new experience. Despite being empirically observed in several problem settings, little…