2 papers
cs.CL2026
Learning Multiple Utterance-Level Attribute Representations with a Unified Speech Encoder
Maryem Bouziane, Salima Mdhaffar, Yannick Estève
Speech foundation models trained with self-supervised learning produce generic speech representations that support a wide range of speech processing tasks. When further adapted wit…
cs.CL2026
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
Phuong-Hang Le, Valentin Pelloin, Arnault Chatelain +27
We release Pantagruel models, a new family of self-supervised encoder models for French text and speech. Instead of predicting modality-tailored targets such as textual tokens or s…