8 papers
Listen to the Features: Voice Anonymization Driven by Content Embedding Matching over Signal Reconstruction
Adrien Schneider, Kacper Zabkowski, Anderson Augusma +3
The paper presents a voice anonymization model focusing on preserving content rather than producing realistic speech. It relies on content embeddings extracted from a frozen pretra…
Anomalies in Multivariate Time Series Benchmarks Are Mostly Univariate
Marc Pinet, Julien Cumin, Samuel Berlemont +1
Many recent multivariate time series anomaly detection (MTSAD) models incorporate cross-channel modeling, under the implicit assumption that the structure of anomalies may be sprea…
Variational Encoder--Multi-Decoder (VE-MD) for Privacy-by-functional-design (Group) Emotion Recognition
Anderson Augusma, Dominique Vaufreydaz, Fédérique Letué
Group Emotion Recognition (GER) aims to infer collective affect in social environments such as classrooms, crowds, and public events. Many existing approaches rely on explicit indi…
MuRAL: A Multi-Resident Ambient Sensor Dataset Annotated with Natural Language for Activities of Daily Living
Xi Chen, Julien Cumin, Fano Ramparany +1
Recent progress in Large Language Models (LLMs) has enabled advanced reasoning and zero-shot recognition for human activity understanding with ambient sensor data. However, widely…
Exploring Dynamic Parameters for Vietnamese Gender-Independent ASR
Sotheara Leang, Ãric Castelli, Dominique Vaufreydaz +1
The dynamic characteristics of speech signal provides temporal information and play an important role in enhancing Automatic Speech Recognition (ASR). In this work, we characterize…
Exploring VQ-VAE with Prosody Parameters for Speaker Anonymization
Sotheara Leang, Anderson Augusma, Eric Castelli +3
Human speech conveys prosody, linguistic content, and speaker identity. This article investigates a novel speaker anonymization approach using an end-to-end network based on a Vect…