28 citations · 29 across the 3 of their papers we have counts for
5 papers
Latin BERT: A Contextual Language Model for Classical Philology
David Bamman, Patrick J. Burns
We present Latin BERT, a contextual language model for the Latin language, trained on 642.7 million words from a variety of sources spanning the Classical era to the 21st century.…
Breaking Speech Recognizers to Imagine Lyrics
Jon Gillick, David Bamman
We introduce a new method for generating text, and in particular song lyrics, based on the speech-like acoustic qualities of a given audio file. We repurpose a vocal source separat…
An Annotated Dataset of Coreference in English Literature
David Bamman, Olivia Lewke, Anya Mansoor
We present in this work a new dataset of coreference annotations for works of literature in English, covering 29,103 mentions in 210,532 tokens from 100 works of fiction. This data…
Learning to Groove with Inverse Sequence Transformations
Jon Gillick, Adam Roberts, Jesse Engel +2
We explore models for translating abstract musical ideas (scores, rhythms) into expressive performances using Seq2Seq and recurrent Variational Information Bottleneck (VIB) models.…
DeepSeek: Content Based Image Search & Retrieval
Tanya Piplani, David Bamman
Most of the internet today is composed of digital media that includes videos and images. With pixels becoming the currency in which most transactions happen on the internet, it is…