Showing eess.ASShow all
2 papers · 1 filter
eess.AS2025
SEAL: Speaker Error Correction using Acoustic-conditioned Large Language Models
Anurag Kumar, Rohit Paturi, Amber Afshan +1
Speaker Diarization (SD) is a crucial component of modern end-to-end ASR pipelines. Traditional SD systems, which are typically audio-based and operate independently of ASR, often…
eess.AS2024
Using RLHF to align speech enhancement approaches to mean-opinion quality scores
Anurag Kumar, Andrew Perrault, Donald S. Williamson
Objective speech quality measures are typically used to assess speech enhancement algorithms, but it has been shown that they are sub-optimal as learning objectives because they do…