3 citations · 3 across the 3 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2025
SEAL: Speaker Error Correction using Acoustic-conditioned Large Language Models
Anurag Kumar, Rohit Paturi, Amber Afshan +1
Speaker Diarization (SD) is a crucial component of modern end-to-end ASR pipelines. Traditional SD systems, which are typically audio-based and operate independently of ASR, often…
eess.AS2022
Curriculum optimization for low-resource speech recognition
Anastasia Kuznetsova, Anurag Kumar, Jennifer Drexler Fox +1
Modern end-to-end speech recognition models show astonishing results in transcribing audio signals into written text. However, conventional data feeding pipelines may be sub-optima…