activity
20192021
most citedA Comparison of Methods for OOV-word Recognition on a New Public Dataset

8 citations · 9 across the 5 of their papers we have counts for

collaborators

10 papers

cs.CL20218 cited

A Comparison of Methods for OOV-word Recognition on a New Public Dataset

Rudolf A. Braun, Srikanth Madikeri, Petr Motlicek

A common problem for automatic speech recognition systems is how to recognize words that they did not see during training. Currently there is no established method of evaluating di…

cs.SD2021

Comparing CTC and LFMMI for out-of-domain adaptation of wav2vec 2.0 acoustic model

Apoorv Vyas, Srikanth Madikeri, Hervé Bourlard

In this work, we investigate if the wav2vec 2.0 self-supervised pretraining helps mitigate the overfitting issues with connectionist temporal classification (CTC) training to reduc…

cs.LG2020

Lattice-Free MMI Adaptation Of Self-Supervised Pretrained Acoustic Models

Apoorv Vyas, Srikanth Madikeri, Hervé Bourlard

In this work, we propose lattice-free MMI (LFMMI) for supervised adaptation of self-supervised pretrained acoustic model. We pretrain a Transformer model on thousand hours of untra…

eess.AS2020

Novel Architectures for Unsupervised Information Bottleneck based Speaker Diarization of Meetings

Nauman Dawalatabad, Srikanth Madikeri, C. Chandra Sekhar +1

Speaker diarization is an important problem that is topical, and is especially useful as a preprocessor for conversational speech related applications. The objective of this paper…

eess.AS2020

Pkwrap: a PyTorch Package for LF-MMI Training of Acoustic Models

Srikanth Madikeri, Sibo Tong, Juan Zuluaga-Gomez +3

We present a simple wrapper that is useful to train acoustic models in PyTorch using Kaldi's LF-MMI training framework. The wrapper, called pkwrap (short form of PyTorch kaldi wrap…

cs.SD2020

Speech Activity Detection Based on Multilingual Speech Recognition System

Seyyed Saeed Sarfjoo, Srikanth Madikeri, Petr Motlicek

To better model the contextual information and increase the generalization ability of Speech Activity Detection (SAD) system, this paper leverages a multi-lingual Automatic Speech…