565 citations · 818 across the 33 of their papers we have counts for
8 papers · 1 filter
Reconsidering the Past: Optimizing Hidden States in Language Models
Davis Yoshida, Kevin Gimpel
We present Hidden-State Optimization (HSO), a gradient-based method for improving the performance of transformer language models at inference time. Similar to dynamic evaluation (K…
Substructure Distribution Projection for Zero-Shot Cross-Lingual Dependency Parsing
Haoyue Shi, Kevin Gimpel, Karen Livescu
We present substructure distribution projection (SubDP), a technique that projects a distribution over structures in one domain to another, by projecting substructure distributions…
On Generalization in Coreference Resolution
Shubham Toshniwal, Patrick Xia, Sam Wiseman +2
While coreference resolution is defined independently of dataset domain, most models for performing coreference resolution do not transfer well to unseen domains. We consolidate a…
TVStoryGen: A Dataset for Generating Stories with Character Descriptions
Mingda Chen, Kevin Gimpel
We introduce TVStoryGen, a story generation dataset that requires generating detailed TV show episode recaps from a brief summary and a set of documents describing the characters i…
Paraphrastic Representations at Scale
John Wieting, Kevin Gimpel, Graham Neubig +1
We present a system that allows users to train their own state-of-the-art paraphrastic sentence representations in a variety of languages. We also release trained models for Englis…
SummScreen: A Dataset for Abstractive Screenplay Summarization
Mingda Chen, Zewei Chu, Sam Wiseman +1
We introduce SummScreen, a summarization dataset comprised of pairs of TV series transcripts and human written recaps. The dataset provides a challenging testbed for abstractive su…