2 papers
cs.LG2026
Spectral Collapse Drives Loss of Plasticity in Deep Continual Learning
Arjun Prakash, Naicheng He, Kaicheng Guo +5
We investigate why deep neural networks suffer from loss of plasticity in continual learning, and thus fail to learn new tasks without reinitializing parameters. We show that this…
cs.LG2026
Bi-Level Policy Optimization with Nyström Hypergradients
Arjun Prakash, Naicheng He, Denizalp Goktas +2
The dependency of the actor on the critic in actor-critic (AC) reinforcement learning means that AC can be characterized as a bilevel optimization (BLO) problem, also called a Stac…