5 citations · 7 across the 4 of their papers we have counts for
3 papers · 1 filter
Two-Pass Low Latency End-to-End Spoken Language Understanding
Siddhant Arora, Siddharth Dalmia, Xuankai Chang +3
End-to-end (E2E) models are becoming increasingly popular for spoken language understanding (SLU) systems and are beginning to achieve competitive performance to pipeline-based app…
Building African Voices
Perez Ogayo, Graham Neubig, Alan W Black
Modern speech synthesis techniques can produce natural-sounding speech given sufficient high-quality data and compute resources. However, such data is not readily available for man…
A Deep Learning Approach to Data-driven Parameterizations for Statistical Parametric Speech Synthesis
Prasanna Kumar Muthukumar, Alan W. Black
Nearly all Statistical Parametric Speech Synthesizers today use Mel Cepstral coefficients as the vocal tract parameterization of the speech signal. Mel Cepstral coefficients were n…