3 papers
eess.AS2026
DAVSS: Distilled Audio-Visual State Space Models
Saurabhchand Bhati, Mrudula Athi, Amit S. Chhetri +1
State-space models (SSMs) distilled from transformer teachers combine the performance of transformers with the efficiency of SSMs. We extend the Transformer-SSM knowledge distillat…
eess.AS2026
USAD 2.0: Scaling Representation Distillation for Universal Audio Understanding
Heng-Jui Chang, Alexander H. Liu, Saurabhchand Bhati +4
Audio encoders are critical to modern audio applications as large language models (LLMs) increasingly rely on a single encoder for diverse inputs. While self-supervised learning (S…
cs.SD2026
Towards Improving Speaker Distance Estimation through Generative Impulse Response Augmentation
Anton Ratnarajah, Mehmet Ergezer, Arun Nair +1
The Room Acoustics and Speaker Distance Estimation (SDE) Challenge at ICASSP 2025 explores the effectiveness of augmented room impulse response (RIR) data for improving SDE model p…