243 citations · 416 across the 4 of their papers we have counts for
6 papers
Unified Scaling Laws for Routed Language Models
Aidan Clark, Diego de las Casas, Aurelia Guy +23
The performance of a language model has been shown to be effectively modeled as a power-law in its parameter count. Here we study the scaling behaviors of Routing Networks: archite…
Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Jack W. Rae, Sebastian Borgeaud, Trevor Cai +77
Language modelling provides a step towards intelligent communication systems by harnessing large repositories of written human knowledge to better predict and understand the world.…
Human-Agent Cooperation in Bridge Bidding
Edward Lockhart, Neil Burch, Nolan Bard +4
We introduce a human-compatible reinforcement-learning approach to a cooperative game, making use of a third-party hand-coded human-compatible bot to generate initial training data…
OpenSpiel: A Framework for Reinforcement Learning in Games
Marc Lanctot, Edward Lockhart, Jean-Baptiste Lespiau +24
OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games. OpenSpiel supports n-player (single- and multi…
Leveraging Sentence Similarity in Natural Language Generation: Improving Beam Search using Range Voting
Sebastian Borgeaud, Guy Emerson
We propose a method for natural language generation, choosing the most representative output rather than the most likely output. By viewing the language generation process from the…
Unsupervised Learning of Object Keypoints for Perception and Control
Tejas Kulkarni, Ankush Gupta, Catalin Ionescu +4
The study of object representations in computer vision has primarily focused on developing representations that are useful for image classification, object detection, or semantic s…