4 papers
Improved low-resource Somali speech recognition by semi-supervised acoustic and language model training
Astik Biswas, Raghav Menon, Ewald van der Westhuizen +1
We present improvements in automatic speech recognition (ASR) for Somali, a currently extremely under-resourced language. This forms part of a continuing United Nations (UN) effort…
Feature exploration for almost zero-resource ASR-free keyword spotting using a multilingual bottleneck extractor and correspondence autoencoders
Raghav Menon, Herman Kamper, Ewald van der Westhuizen +2
We compare features for dynamic time warping (DTW) when used to bootstrap keyword spotting (KWS) in an almost zero-resource setting. Such quickly-deployable systems aim to support…
Automatic Speech Recognition for Humanitarian Applications in Somali
Raghav Menon, Astik Biswas, Armin Saeb +2
We present our first efforts in building an automatic speech recognition system for Somali, an under-resourced language, using 1.57 hrs of annotated speech for acoustic model train…
Fast ASR-free and almost zero-resource keyword spotting using DTW and CNNs for humanitarian monitoring
Raghav Menon, Herman Kamper, John Quinn +1
We use dynamic time warping (DTW) as supervision for training a convolutional neural network (CNN) based keyword spotting system using a small set of spoken isolated keywords. The…