5 citations · 6 across the 4 of their papers we have counts for
7 papers · 1 filter
LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech
Titouan Parcollet, Ha Nguyen, Solene Evain +19
Self-supervised learning (SSL) is at the origin of unprecedented improvements in many different domains including computer vision and natural language processing. Speech processing…
FlauBERT: Unsupervised Language Model Pre-training for French
Hang Le, Loïc Vial, Jibril Frej +7
Language models have become a key step to achieve state-of-the art results in many different Natural Language Processing (NLP) tasks. Leveraging the huge amount of unlabeled texts…
Empirical Study of Diachronic Word Embeddings for Scarce Data
Syrielle Montariol, Alexandre Allauzen
Word meaning change can be inferred from drifts of time-varying word embeddings. However, temporal data may be too sparse to build robust word embeddings and to discriminate signif…
Learning dynamic word embeddings with drift regularisation
Syrielle Montariol, Alexandre Allauzen
Word usage, meaning and connotation change throughout time. Diachronic word embeddings are used to grasp these changes in an unsupervised way. In this paper, we use variants of the…
Exploring sentence informativeness
Syrielle Montariol, Aina Garí Soler, Alexandre Allauzen
This study is a preliminary exploration of the concept of informativeness -how much information a sentence gives about a word it contains- and its potential benefits to building qu…
Word Usage Similarity Estimation with Sentence Representations and Automatic Substitutes
Aina Garí Soler, Marianna Apidianaki, Alexandre Allauzen
Usage similarity estimation addresses the semantic proximity of word instances in different contexts. We apply contextualized (ELMo and BERT) word and sentence embeddings to this t…