3 citations · 6 across the 9 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2025
Audiobook-CC: Controllable Long-context Speech Generation for Multicast Audiobook
Min Liu, JingJing Yin, Xiang Zhang +6
Existing text-to-speech systems predominantly focus on single-sentence synthesis and lack adequate contextual modeling as well as fine-grained performance control capabilities for…
eess.AS2023
PP-MeT: a Real-world Personalized Prompt based Meeting Transcription System
Xiang Lyu, Yuhang Cao, Qing Wang +5
Speaker-attributed automatic speech recognition (SA-ASR) improves the accuracy and applicability of multi-speaker ASR systems in real-world scenarios by assigning speaker labels to…
eess.AS2023★ 1 cited
PromptVC: Flexible Stylistic Voice Conversion in Latent Space Driven by Natural Language Prompts
Jixun Yao, Yuguang Yang, Yi Lei +7
Style voice conversion aims to transform the style of source speech to a desired style according to real-world application demands. However, the current style voice conversion appr…