3 papers
cs.CL2026
Speaker Group Encoding in Self-supervised Speech Recognition Models
Felix Herron, Solange Rossato Alexandre Allauzen, Benoit Favre +1
We investigate what self-supervised speech recognition models (S3Ms) learn about speaker groups (SGs). We examine several states of S3Ms: pretrained, finetuned on speaker identific…
cs.CL2026
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
Phuong-Hang Le, Valentin Pelloin, Arnault Chatelain +27
We release Pantagruel models, a new family of self-supervised encoder models for French text and speech. Instead of predicting modality-tailored targets such as textual tokens or s…
cs.HC2024
THERADIA WoZ: An Ecological Corpus for Appraisal-based Affect Research in Healthcare
Hippolyte Fournier, Sina Alisamir, Safaa Azzakhnini +14
We present THERADIA WoZ, an ecological corpus designed for audiovisual research on affect in healthcare. Two groups of senior individuals, consisting of 52 healthy participants and…