182 citations · 879 across the 27 of their papers we have counts for
13 papers · 1 filter
RecurrentGemma: Moving Past Transformers for Efficient Open Language Models
Aleksandar Botev, Soham De, Samuel L Smith +59
We introduce RecurrentGemma, a family of open language models which uses Google's novel Griffin architecture. Griffin combines linear recurrences with local attention to achieve ex…
Hindering Adversarial Attacks with Implicit Neural Representations
Andrei A. Rusu, Dan A. Calian, Sven Gowal +1
We introduce the Lossy Implicit Network Activation Coding (LINAC) defence, an input transformation which successfully hinders several common adversarial attacks on CIFAR- class…
Probing Transfer in Deep Reinforcement Learning without Task Engineering
Andrei A. Rusu, Sebastian Flennerhag, Dushyant Rao +2
We evaluate the use of original game curricula supported by the Atari 2600 console as a heterogeneous transfer benchmark for deep reinforcement learning agents. Game designers crea…
MO2: Model-Based Offline Options
Sasha Salter, Markus Wulfmeier, Dhruva Tirumala +4
The ability to discover useful behaviours from past experience and transfer them to new tasks is considered a core component of natural embodied intelligence. Inspired by neuroscie…
Skillful Precipitation Nowcasting using Deep Generative Models of Radar
Suman Ravuri, Karel Lenc, Matthew Willson +17
Precipitation nowcasting, the high-resolution forecasting of precipitation up to two hours ahead, supports the real-world socio-economic needs of many sectors reliant on weather-de…
A Distributional View on Multi-Objective Policy Optimization
Abbas Abdolmaleki, Sandy H. Huang, Leonard Hasenclever +7
Many real-world problems require trading off multiple competing objectives. However, these objectives are often in different units and/or scales, which can make it challenging for…