3 papers
cs.CL2026
ALARM: Audio-Language Alignment for Reasoning Models
Petr Grinberg, Hassan Shahmohammadi
Large audio language models (ALMs) extend LLMs with auditory understanding. A common approach freezes the LLM and trains only an adapter on self-generated targets. However, this fa…
eess.AS2025
A Data-Driven Diffusion-based Approach for Audio Deepfake Explanations
Petr Grinberg, Ankur Kumar, Surya Koppisetti +1
Evaluating explainability techniques, such as SHAP and LRP, in the context of audio deepfake detection is challenging due to lack of clear ground truth annotations. In the cases wh…
cs.LG2025
What Does an Audio Deepfake Detector Focus on? A Study in the Time Domain
Petr Grinberg, Ankur Kumar, Surya Koppisetti +1
Adding explanations to audio deepfake detection (ADD) models will boost their real-world application by providing insight on the decision making process. In this paper, we propose…