50 citations · 51 across the 3 of their papers we have counts for
3 papers
eess.AS2024★ 1 cited
Parameter-Efficient Transfer Learning under Federated Learning for Automatic Speech Recognition
Xuan Kan, Yonghui Xiao, Tien-Ju Yang +2
This work explores the challenge of enhancing Automatic Speech Recognition (ASR) model performance across various user-specific domains while preserving user data privacy. We emplo…
cs.CL2024
Text Injection for Neural Contextual Biasing
Zhong Meng, Zelin Wu, Rohit Prabhavalkar +5
Neural contextual biasing effectively improves automatic speech recognition (ASR) for crucial phrases within a speaker's context, particularly those that are infrequent in the trai…
cs.SD2023★ 50 cited
Noise2Music: Text-conditioned Music Generation with Diffusion Models
Qingqing Huang, Daniel S. Park, Tao Wang +12
We introduce Noise2Music, where a series of diffusion models is trained to generate high-quality 30-second music clips from text prompts. Two types of diffusion models, a generator…