From the 1 of 5 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Deep Reinforcement Learning and The Tale of Two Temporal Difference Errors
Juan Sebastian Rojas, Chi-Guhn Lee
The temporal difference (TD) error was first formalized in Sutton (1988), where it was first characterized as the difference between temporally successive predictions, and later, i…
cs.LG2026
Thermodynamics of Reinforcement Learning Curricula
Jacob Adamczyk, Juan Sebastian Rojas, Rahul V. Kulkarni
Connections between statistical mechanics and machine learning have repeatedly proven fruitful, providing insight into optimization, generalization, and representation learning. In…