3.8k citations · 4.1k across the 13 of their papers we have counts for
Showing 2017 · cs.CLShow all
2 papers · 2 filters
cs.CL2017
Advances in Pre-Training Distributed Word Representations
Tomas Mikolov, Edouard Grave, Piotr Bojanowski +2
Many Natural Language Processing applications nowadays rely on pre-trained word representations estimated from large text corpora such as news collections, Wikipedia and Web Crawl.…
cs.CL2017
Learning Simpler Language Models with the Differential State Framework
Alexander G. Ororbia, Tomas Mikolov, David Reitter
Learning useful information across long time lags is a critical and difficult problem for temporal neural models in tasks such as language modeling. Existing architectures that add…