4 citations · 7 across the 6 of their papers we have counts for
6 papers
Generalized Multilingual Text-to-Speech Generation with Language-Aware Style Adaptation
Haowei Lou, Hye-young Paik, Sheng Li +2
Text-to-Speech (TTS) models can generate natural, human-like speech across multiple languages by transforming phonemes into waveforms. However, multilingual TTS remains challenging…
LatentSpeech: Latent Diffusion for Text-To-Speech Generation
Haowei Lou, Helen Paik, Pari Delir Haghighi +2
Diffusion-based Generative AI gains significant attention for its superior performance over other generative techniques like Generative Adversarial Networks and Variational Autoenc…
Aligner-Guided Training Paradigm: Advancing Text-to-Speech Models with Aligner Guided Duration
Haowei Lou, Helen Paik, Wen Hu +1
Recent advancements in text-to-speech (TTS) systems, such as FastSpeech and StyleSpeech, have significantly improved speech generation quality. However, these models often rely on…
Designing Efficient LLM Accelerators for Edge Devices
Jude Haris, Rappy Saha, Wenhao Hu +1
The increase in open-source availability of Large Language Models (LLMs) has enabled users to deploy them on more and more resource-constrained edge devices to reduce reliance on n…
CityCraft: A Real Crafter for 3D City Generation
Jie Deng, Wenhao Chai, Junsheng Huang +9
City scene generation has gained significant attention in autonomous driving, smart city development, and traffic simulation. It helps enhance infrastructure planning and monitorin…
Deep Learning Methods for Small Molecule Drug Discovery: A Survey
Wenhao Hu, Yingying Liu, Xuanyu Chen +4
With the development of computer-assisted techniques, research communities including biochemistry and deep learning have been devoted into the drug discovery field for over a decad…