8 citations · 9 across the 5 of their papers we have counts for
10 papers
A Comparison of Methods for OOV-word Recognition on a New Public Dataset
Rudolf A. Braun, Srikanth Madikeri, Petr Motlicek
A common problem for automatic speech recognition systems is how to recognize words that they did not see during training. Currently there is no established method of evaluating di…
Comparing CTC and LFMMI for out-of-domain adaptation of wav2vec 2.0 acoustic model
Apoorv Vyas, Srikanth Madikeri, Hervé Bourlard
In this work, we investigate if the wav2vec 2.0 self-supervised pretraining helps mitigate the overfitting issues with connectionist temporal classification (CTC) training to reduc…
Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models
Apoorv Vyas, Srikanth Madikeri, Hervé Bourlard
In this work, we propose lattice-free MMI (LFMMI) for supervised adaptation of self-supervised pretrained acoustic model. We pretrain a Transformer model on thousand hours of untra…
Novel Architectures for Unsupervised Information Bottleneck based Speaker Diarization of Meetings
Nauman Dawalatabad, Srikanth Madikeri, C. Chandra Sekhar +1
Speaker diarization is an important problem that is topical, and is especially useful as a preprocessor for conversational speech related applications. The objective of this paper…
Pkwrap: a PyTorch Package for LF-MMI Training of Acoustic Models
Srikanth Madikeri, Sibo Tong, Juan Zuluaga-Gomez +3
We present a simple wrapper that is useful to train acoustic models in PyTorch using Kaldi's LF-MMI training framework. The wrapper, called pkwrap (short form of PyTorch kaldi wrap…
Speech Activity Detection Based on Multilingual Speech Recognition System
Seyyed Saeed Sarfjoo, Srikanth Madikeri, Petr Motlicek
To better model the contextual information and increase the generalization ability of Speech Activity Detection (SAD) system, this paper leverages a multi-lingual Automatic Speech…