1 paper
Haodong Liang, Lifeng Lai
We investigate the ability of transformers to perform in-context reinforcement learning (ICRL), where a model must infer and execute learning algorithms from trajectory data withou…