3 papers
cs.SD2025
Audio Question Answering with GRPO-Based Fine-Tuning and Calibrated Segment-Level Predictions
Marcel Gibier, Nolwenn Celton, Raphaël Duroselle +3
In this report, we describe our submission to Track 5 of the DCASE 2025 Challenge for the task of Audio Question Answering(AQA). Our system leverages the SSL backbone BEATs to extr…
cs.SD2025
Segmentwise Pruning in Audio-Language Models
Marcel Gibier, Raphaël Duroselle, Pierre Serrano +2
Recent audio-language models have shown impressive performance across a wide range of audio tasks and are increasingly capable of handling long audio inputs. However, the computing…
cs.SD2025
Improving Out-of-Domain Audio Deepfake Detection via Layer Selection and Fusion of SSL-Based Countermeasures
Pierre Serrano, Raphaël Duroselle, Florian Angulo +2
Audio deepfake detection systems based on frozen pre-trained self-supervised learning (SSL) encoders show a high level of performance when combined with layer-weighted pooling meth…