activity
20152022
most citedTransition-Based Dependency Parsing with Stack Long Short-Term Memory

526 citations · 1.2k across the 20 of their papers we have counts for

collaborators
Showing cs.CLShow all

35 papers · 1 filter

cs.CL2022

Enabling arbitrary translation objectives with Adaptive Tree Search

Wang Ling, Wojciech Stokowiec, Domenic Donato +4

We introduce an adaptive tree search algorithm, that can find high-scoring outputs under translation models that make no assumptions about the form or structure of the search objec…

cs.CL2022243 cited

Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Jack W. Rae, Sebastian Borgeaud, Trevor Cai +77

Language modelling provides a step towards intelligent communication systems by harnessing large repositories of written human knowledge to better predict and understand the world.…

cs.CL20217 cited

Diverse Pretrained Context Encodings Improve Document Translation

Domenic Donato, Lei Yu, Chris Dyer

We propose a new architecture for adapting a sentence-level sequence-to-sequence transformer by incorporating multiple pretrained document context signals and assess the impact on…

cs.CL2020

Syntactic Structure Distillation Pretraining For Bidirectional Encoders

Adhiguna Kuncoro, Lingpeng Kong, Daniel Fried +4

Textual representation learners trained on large amounts of data have achieved notable success on downstream tasks; intriguingly, they have also performed well on challenging tests…

cs.CL20203 cited

Learning to Segment Actions from Observation and Narration

Daniel Fried, Jean-Baptiste Alayrac, Phil Blunsom +3

We apply a generative segmental model of task structure, guided by narration, to action segmentation in video. We focus on unsupervised and weakly-supervised settings where no acti…

cs.CL202013 cited

Learning Robust and Multilingual Speech Representations

Kazuya Kawakami, Luyu Wang, Chris Dyer +2

Unsupervised speech representation learning has shown remarkable success at finding representations that correlate with phonetic structures and improve downstream speech recognitio…