5 papers · 1 filter
Speaker Group Encoding in Self-supervised Speech Recognition Models
Felix Herron, Solange Rossato Alexandre Allauzen, Benoit Favre +1
We investigate what self-supervised speech recognition models (S3Ms) learn about speaker groups (SGs). We examine several states of S3Ms: pretrained, finetuned on speaker identific…
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
Phuong-Hang Le, Valentin Pelloin, Arnault Chatelain +27
We release Pantagruel models, a new family of self-supervised encoder models for French text and speech. Instead of predicting modality-tailored targets such as textual tokens or s…
LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech
Titouan Parcollet, Ha Nguyen, Solene Evain +19
Self-supervised learning (SSL) is at the origin of unprecedented improvements in many different domains including computer vision and natural language processing. Speech processing…
Gender Representation in Open Source Speech Resources
Mahault Garnerin, Solange Rossato, Laurent Besacier
With the rise of artificial intelligence (AI) and the growing use of deep-learning architectures, the question of ethics, transparency and fairness of AI systems has become a centr…
Gender Representation in French Broadcast Corpora and Its Impact on ASR Performance
Mahault Garnerin, Solange Rossato, Laurent Besacier
This paper analyzes the gender representation in four major corpora of French broadcast. These corpora being widely used within the speech processing community, they are a primary…