20 citations · 41 across the 22 of their papers we have counts for
4 papers · 1 filter
Domain Adaptation for Contrastive Audio-Language Models
Soham Deshmukh, Rita Singh, Bhiksha Raj
Audio-Language Models (ALM) aim to be general-purpose audio models by providing zero-shot capabilities at test time. The zero-shot performance of ALM improves by using suitable tex…
uSee: Unified Speech Enhancement and Editing with Conditional Diffusion Models
Muqiao Yang, Chunlei Zhang, Yong Xu +4
Speech enhancement aims to improve the quality of speech signals in terms of quality and intelligibility, and speech editing refers to the process of editing the speech according t…
Psychoacoustic Challenges Of Speech Enhancement On VoIP Platforms
Joseph Konan, Shikhar Agnihotri, Ojas Bhargave +4
Within the ambit of VoIP (Voice over Internet Protocol) telecommunications, the complexities introduced by acoustic transformations merit rigorous analysis. This research, rooted i…
Prompting Audios Using Acoustic Properties For Emotion Representation
Hira Dhamyal, Benjamin Elizalde, Soham Deshmukh +3
Emotions lie on a continuum, but current models treat emotions as a finite valued discrete variable. This representation does not capture the diversity in the expression of emotion…