activity
20142024
most citedRecurrent Models of Visual Attention

999 citations · 1.2k across the 16 of their papers we have counts for

collaborators
Showing cs.LGShow all

7 papers · 1 filter

cs.LG202412 cited

Genie: Generative Interactive Environments

Jake Bruce, Michael Dennis, Ashley Edwards +22

We introduce Genie, the first generative interactive environment trained in an unsupervised manner from unlabelled Internet videos. The model can be prompted to generate an endless…

cs.LG20242 cited

Offline Actor-Critic Reinforcement Learning Scales to Large Models

Jost Tobias Springenberg, Abbas Abdolmaleki, Jingwei Zhang +9

We show that offline actor-critic reinforcement learning can scale to large models - such as transformers - and follows similar scaling laws as supervised learning. We find that of…

cs.LG20231 cited

TacticAI: an AI assistant for football tactics

Zhe Wang, Petar Veličković, Daniel Hennes +20

Identifying key patterns of tactics implemented by rival teams, and developing effective responses, lies at the heart of modern football. However, doing so algorithmically remains…

cs.LG2023

Policy composition in reinforcement learning via multi-objective policy optimization

Shruti Mishra, Ankit Anand, Jordan Hoffmann +4

We enable reinforcement learning agents to learn successful behavior policies by utilizing relevant pre-existing teacher policies. The teacher policies are introduced as objectives…

cs.LG20233 cited

Lossless Adaptation of Pretrained Vision Models For Robotic Manipulation

Mohit Sharma, Claudio Fantacci, Yuxiang Zhou +4

Recent works have shown that large models pretrained on common visual learning tasks can provide useful representations for a wide range of specialized perception problems, as well…

cs.LG20226 cited

Retrieval-Augmented Reinforcement Learning

Anirudh Goyal, Abram L. Friesen, Andrea Banino +13

Most deep reinforcement learning (RL) algorithms distill experience into parametric behavior policies or value functions via gradient updates. While effective, this approach has se…