1.6k citations · 1.7k across the 2 of their papers we have counts for
2 papers
cs.LG2019★ 46 cited
Augmenting Self-attention with Persistent Memory
Sainbayar Sukhbaatar, Edouard Grave, Guillaume Lample +2
Transformer networks have lead to important progress in language modeling and machine translation. These models include two consecutive modules, a feed-forward layer and a self-att…
cs.CL2019★ 1.6k cited
Cross-lingual Language Model Pretraining
Guillaume Lample, Alexis Conneau
Recent studies have demonstrated the efficiency of generative pretraining for English natural language understanding. In this work, we extend this approach to multiple languages an…