activity
20172024
most citedGemini 1.5: Unlocking multimodal understanding across millions of tokens of context

297 citations · 436 across the 12 of their papers we have counts for

collaborators
Showing 2020 · cs.CLShow all

5 papers · 2 filters

cs.CL2020

mT5: A massively multilingual pre-trained text-to-text transformer

Linting Xue, Noah Constant, Adam Roberts +5

The recent "Text-to-Text Transfer Transformer" (T5) leveraged a unified text-to-text format and scale to attain state-of-the-art results on a wide variety of English-language NLP t…

cs.CL2020

Explicit Alignment Objectives for Multilingual Bidirectional Encoders

Junjie Hu, Melvin Johnson, Orhan Firat +2

Pre-trained cross-lingual encoders such as mBERT (Devlin et al., 2019) and XLMR (Conneau et al., 2020) have proven to be impressively effective at enabling transfer-learning of NLP…

cs.CL2020

Harnessing Multilinguality in Unsupervised Machine Translation for Rare Languages

Xavier Garcia, Aditya Siddhant, Orhan Firat +1

Unsupervised translation has reached impressive performance on resource-rich language pairs such as English-French and English-German. However, early studies have shown that in mor…

cs.CL2020★ 35 cited

Leveraging Monolingual Data with Self-Supervision for Multilingual Neural Machine Translation

Aditya Siddhant, Ankur Bapna, Yuan Cao +5

Over the last few years two promising research directions in low-resource neural machine translation (NMT) have emerged. The first focuses on utilizing high-resource languages to i…

cs.CL2020

XTREME: A Massively Multilingual Multi-task Benchmark for Evaluating Cross-lingual Generalization

Junjie Hu, Sebastian Ruder, Aditya Siddhant +3

Much recent progress in applications of machine learning models to NLP has been driven by benchmarks that evaluate models across a wide variety of tasks. However, these broad-cover…