1 citations · 2 across the 4 of their papers we have counts for
4 papers
VowelPrompt: Hearing Speech Emotions from Text via Vowel-level Prosodic Augmentation
Yancheng Wang, Osama Hanna, Ruiming Xie +11
Emotion recognition in speech presents a complex multimodal challenge, requiring comprehension of both linguistic content and vocal expressivity, particularly prosodic features suc…
A Domain Adaptation Framework for Speech Recognition Systems with Only Synthetic data
Minh Tran, Yutong Pang, Debjyoti Paul +7
We introduce DAS (Domain Adaptation with Synthetic data), a novel domain adaptation framework for pre-trained ASR model, designed to efficiently adapt to various language-defined d…
LLaMA based Punctuation Restoration With Forward Pass Only Decoding
Yutong Pang, Debjyoti Paul, Kevin Jiang +2
This paper introduces two advancements in the field of Large Language Model Annotation with a focus on punctuation restoration tasks. Our first contribution is the application of L…
Towards scalable efficient on-device ASR with transfer learning
Laxmi Pandey, Ke Li, Jinxi Guo +4
Multilingual pretraining for transfer learning significantly boosts the robustness of low-resource monolingual ASR models. This study systematically investigates three main aspects…