activity
20162024
most citedBeyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

565 citations · 818 across the 33 of their papers we have counts for

collaborators
Showing 2021Show all

8 papers · 1 filter

cs.CL2021

Reconsidering the Past: Optimizing Hidden States in Language Models

Davis Yoshida, Kevin Gimpel

We present Hidden-State Optimization (HSO), a gradient-based method for improving the performance of transformer language models at inference time. Similar to dynamic evaluation (K…

cs.CL2021

Substructure Distribution Projection for Zero-Shot Cross-Lingual Dependency Parsing

Haoyue Shi, Kevin Gimpel, Karen Livescu

We present substructure distribution projection (SubDP), a technique that projects a distribution over structures in one domain to another, by projecting substructure distributions…

cs.CL2021

On Generalization in Coreference Resolution

Shubham Toshniwal, Patrick Xia, Sam Wiseman +2

While coreference resolution is defined independently of dataset domain, most models for performing coreference resolution do not transfer well to unseen domains. We consolidate a…

cs.CL2021

TVStoryGen: A Dataset for Generating Stories with Character Descriptions

Mingda Chen, Kevin Gimpel

We introduce TVStoryGen, a story generation dataset that requires generating detailed TV show episode recaps from a brief summary and a set of documents describing the characters i…

cs.CL2021

Paraphrastic Representations at Scale

John Wieting, Kevin Gimpel, Graham Neubig +1

We present a system that allows users to train their own state-of-the-art paraphrastic sentence representations in a variety of languages. We also release trained models for Englis…

cs.CL2021

SummScreen: A Dataset for Abstractive Screenplay Summarization

Mingda Chen, Zewei Chu, Sam Wiseman +1

We introduce SummScreen, a summarization dataset comprised of pairs of TV series transcripts and human written recaps. The dataset provides a challenging testbed for abstractive su…