2 papers
cs.SD2025
Sparse Autoencoders Make Audio Foundation Models more Explainable
Théo Mariotte, Martin Lebourdais, Antonio Almudévar +3
Audio pretrained models are widely employed to solve various tasks in speech processing, sound event detection, or music information retrieval. However, the representations learned…
cs.SD2025
Sparse deepfake detection promotes better disentanglement
Antoine Teissier, Marie Tahon, Nicolas Dugué +1
Due to the rapid progress of speech synthesis, deepfake detection has become a major concern in the speech processing community. Because it is a critical task, systems must not onl…