2 papers
eess.AS2024
Real-Time and Accurate: Zero-shot High-Fidelity Singing Voice Conversion with Multi-Condition Flow Synthesis
Hui Li, Hongyu Wang, Zhijin Chen +2
Singing voice conversion is to convert the source singing voice into the target singing voice except for the content. Currently, flow-based models can complete the task of voice co…
eess.AS2024
A New Perspective on Speaker Verification: Joint Modeling with DFSMN and Transformer
Hongyu Wang, Hui Li, Bo Li
Speaker verification is to judge the similarity between two unknown voices in an open set, where the ideal speaker embedding should be able to condense discriminant information int…