32 citations · 32 across the 3 of their papers we have counts for
3 papers
cs.CL2022
Few-shot Mining of Naturally Occurring Inputs and Outputs
Mandar Joshi, Terra Blevins, Mike Lewis +2
Creating labeled natural language training data is expensive and requires significant human effort. We mine input output examples from large corpora using a supervised mining funct…
cs.CL2021★ 32 cited
HTLM: Hyper-Text Pre-Training and Prompting of Language Models
Armen Aghajanyan, Dmytro Okhonko, Mike Lewis +4
We introduce HTLM, a hyper-text language model trained on a large-scale web crawl. Modeling hyper-text has a number of advantages: (1) it is easily gathered at scale, (2) it provid…
cs.CL2021
DESCGEN: A Distantly Supervised Dataset for Generating Abstractive Entity Descriptions
Weijia Shi, Mandar Joshi, Luke Zettlemoyer
Short textual descriptions of entities provide summaries of their key attributes and have been shown to be useful sources of background knowledge for tasks such as entity linking a…