22 citations · 22 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 22 cited
Human-Timescale Adaptation in an Open-Ended Task Space
Adaptive Agent Team, Jakob Bauer, Kate Baumli +25
Foundation models have shown impressive adaptation and scalability in supervised and self-supervised learning problems, but so far these successes have not fully translated to rein…
cs.LG2022
An Empirical Study of Implicit Regularization in Deep Offline RL
Caglar Gulcehre, Srivatsan Srinivasan, Jakub Sygnowski +5
Deep neural networks are the most commonly used function approximators in offline reinforcement learning. Prior works have shown that neural nets trained with TD-learning and gradi…