7 citations · 7 across the 5 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
eess.AS2022
Preserving background sound in noise-robust voice conversion via multi-task learning
Jixun Yao, Yi Lei, Qing Wang +6
Background sound is an informative form of art that is helpful in providing a more immersive experience in real-application voice conversion (VC) scenarios. However, prior research…
eess.AS2022★ 7 cited
IQDUBBING: Prosody modeling based on discrete self-supervised speech representation for expressive voice conversion
Wendong Gan, Bolong Wen, Ying Yan +6
Prosody modeling is important, but still challenging in expressive voice conversion. As prosody is difficult to model, and other factors, e.g., speaker, environment and content, wh…