2 papers
cs.AI2026
Learning the Preferences of a Learning Agent
Karim Abdel Sadek, Mark Bedaywi, Rhys Gould +1
For AI systems to be useful to humans, they must understand and act in accordance with our values and preferences. Since specifying preferences is a hard task, inverse reinforcemen…
cs.LG2024
Continuous-Time Analysis of Adaptive Optimization and Normalization
Rhys Gould, Hidenori Tanaka
Adaptive optimization algorithms, particularly Adam and its variant AdamW, are fundamental components of modern deep learning. However, their training dynamics lack comprehensive t…