28 citations · 44 across the 8 of their papers we have counts for
Showing 2022 · cs.LGShow all
2 papers · 2 filters
cs.LG2022★ 10 cited
In-context Reinforcement Learning with Algorithm Distillation
Michael Laskin, Luyu Wang, Junhyuk Oh +11
We propose Algorithm Distillation (AD), a method for distilling reinforcement learning (RL) algorithms into neural networks by modeling their training histories with a causal seque…
cs.LG2022
Uniqueness and Complexity of Inverse MDP Models
Marcus Hutter, Steven Hansen
What is the action sequence aa'a" that was likely responsible for reaching state s"' (from state s) in 3 steps? Addressing such questions is important in causal reasoning and in re…