Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
When do spectral gradient updates help in deep learning?
Damek Davis, Dmitriy Drusvyatskiy
Spectral gradient methods, such as the recently popularized Muon optimizer, are a promising alternative to standard Euclidean gradient descent for training deep neural networks and…
cs.LG2025
What is the objective of reasoning with reinforcement learning?
Damek Davis, Benjamin Recht
We show that several popular algorithms for reinforcement learning in large language models with binary rewards can be viewed as stochastic gradient ascent on a monotone transform…