1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CL2025
CoLMbo: Speaker Language Model for Descriptive Profiling
Massa Baali, Shuo Han, Syed Abdul Hannan +5
Speaker recognition systems are often limited to classification tasks and struggle to generate detailed speaker characteristics or provide context-rich descriptions. These models p…
cs.SD2025
ADIFF: Explaining audio difference using natural language
Soham Deshmukh, Shuo Han, Rita Singh +1
Understanding and explaining differences between audio recordings is crucial for fields like audio forensics, quality assessment, and audio generation. This involves identifying an…
eess.AS2024★ 1 cited
DeWinder: Single-Channel Wind Noise Reduction using Ultrasound Sensing
Kuang Yuan, Shuo Han, Swarun Kumar +1
The quality of audio recordings in outdoor environments is often degraded by the presence of wind. Mitigating the impact of wind noise on the perceptual quality of single-channel s…