26 citations · 85 across the 16 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2022
Non-Parallel Voice Conversion for ASR Augmentation
Gary Wang, Andrew Rosenberg, Bhuvana Ramabhadran +4
Automatic speech recognition (ASR) needs to be robust to speaker differences. Voice Conversion (VC) modifies speaker characteristics of input speech. This is an attractive feature…
cs.SD2022★ 1 cited
Ask2Mask: Guided Data Selection for Masked Speech Modeling
Murali Karthick Baskar, Andrew Rosenberg, Bhuvana Ramabhadran +2
Masked speech modeling (MSM) methods such as wav2vec2 or w2v-BERT learn representations over speech frames which are randomly masked within an utterance. While these methods improv…