most citedImproving accuracy of rare words for RNN-Transducer through unigram shallow fusion

7 citations · 11 across the 5 of their papers we have counts for

collaborators

5 papers

eess.AS20221 cited

Automatic Dialect Density Estimation for African American English

Alexander Johnson, Kevin Everson, Vijay Ravi +3

In this paper, we explore automatic prediction of dialect density of the African American English (AAE) dialect, where dialect density is defined as the percentage of words in an u…

eess.AS20222 cited

FrAUG: A Frame Rate Based Data Augmentation Method for Depression Detection from Speech Signals

Vijay Ravi, Jinhan Wang, Jonathan Flint +1

In this paper, a data augmentation method is proposed for depression detection from speech signals. Samples for data augmentation were created by changing the frame-width and the f…

cs.CL20207 cited

Improving accuracy of rare words for RNN-Transducer through unigram shallow fusion

Vijay Ravi, Yile Gu, Ankur Gandhe +5

End-to-end automatic speech recognition (ASR) systems, such as recurrent neural network transducer (RNN-T), have become popular, but rare word remains a challenge. In this paper, w…

eess.AS20201 cited

Variable frame rate-based data augmentation to handle speaking-style variability for automatic speaker verification

Amber Afshan, Jinxi Guo, Soo Jin Park +3

The effects of speaking-style variability on automatic speaker verification were investigated using the UCLA Speaker Variability database which comprises multiple speaking styles p…

eess.AS2020

Exploring the Use of an Unsupervised Autoregressive Model as a Shared Encoder for Text-Dependent Speaker Verification

Vijay Ravi, Ruchao Fan, Amber Afshan +2

In this paper, we propose a novel way of addressing text-dependent automatic speaker verification (TD-ASV) by using a shared-encoder with task-specific decoders. An autoregressive…