67 citations · 112 across the 7 of their papers we have counts for
9 papers · 1 filter
Towards Universality in Multilingual Text Rewriting
Xavier Garcia, Noah Constant, Mandy Guo +1
In this work, we take the first steps towards building a universal rewriter: a model capable of rewriting text in any language to exhibit a wide variety of attributes, including st…
Neural Retrieval for Question Answering with Cross-Attention Supervised Data Augmentation
Yinfei Yang, Ning Jin, Kuo Lin +2
Neural models that independently project questions and answers into a shared embedding space allow for efficient continuous space retrieval from large corpora. Independently comput…
MultiReQA: A Cross-Domain Evaluation for Retrieval Question Answering Models
Mandy Guo, Yinfei Yang, Daniel Cer +2
Retrieval question answering (ReQA) is the task of retrieving a sentence-level answer to a question from an open corpus (Ahmad et al.,2019).This paper presents MultiReQA, anew mult…
Bridging the Gap for Tokenizer-Free Language Models
Dokook Choe, Rami Al-Rfou, Mandy Guo +2
Purely character-based language models (LMs) have been lagging in quality on large scale datasets, and current state-of-the-art LMs rely on word tokenization. It has been assumed t…
Multilingual Universal Sentence Encoder for Semantic Retrieval
Yinfei Yang, Daniel Cer, Amin Ahmad +9
We introduce two pre-trained retrieval focused multilingual sentence encoding models, respectively based on the Transformer and CNN model architectures. The models embed text from…
Hierarchical Document Encoder for Parallel Corpus Mining
Mandy Guo, Yinfei Yang, Keith Stevens +5
We explore using multilingual document embeddings for nearest neighbor mining of parallel data. Three document-level representations are investigated: (i) document embeddings gener…