activity
20192026
collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2026

Speaker Group Encoding in Self-supervised Speech Recognition Models

Felix Herron, Solange Rossato Alexandre Allauzen, Benoit Favre +1

We investigate what self-supervised speech recognition models (S3Ms) learn about speaker groups (SGs). We examine several states of S3Ms: pretrained, finetuned on speaker identific…

cs.CL2026

Pantagruel: Unified Self-Supervised Encoders for French Text and Speech

Phuong-Hang Le, Valentin Pelloin, Arnault Chatelain +27

We release Pantagruel models, a new family of self-supervised encoder models for French text and speech. Instead of predicting modality-tailored targets such as textual tokens or s…

cs.CL2023

LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech

Titouan Parcollet, Ha Nguyen, Solene Evain +19

Self-supervised learning (SSL) is at the origin of unprecedented improvements in many different domains including computer vision and natural language processing. Speech processing…

cs.CL2020

Gender Representation in Open Source Speech Resources

Mahault Garnerin, Solange Rossato, Laurent Besacier

With the rise of artificial intelligence (AI) and the growing use of deep-learning architectures, the question of ethics, transparency and fairness of AI systems has become a centr…

cs.CL2019

Gender Representation in French Broadcast Corpora and Its Impact on ASR Performance

Mahault Garnerin, Solange Rossato, Laurent Besacier

This paper analyzes the gender representation in four major corpora of French broadcast. These corpora being widely used within the speech processing community, they are a primary…