3 papers
eess.AS2025
A Comprehensive Survey on Generative AI for Video-to-Music Generation
Shulei Ji, Songruoyao Wu, Zihao Wang +2
The burgeoning growth of video-to-music generation can be attributed to the ascendancy of multimodal generative models. However, there is a lack of literature that comprehensively…
eess.AS2025
Assessing Data Replication in Symbolic Music via Adapted Structural Similarity Index Measure
Shulei Ji, Zihao Wang, Le Ma +2
AI-generated music may inadvertently replicate samples from the training data, raising concerns of plagiarism. Similarity measures can quantify such replication, thereby offering s…
cs.SD2024
SaMoye: Zero-shot Singing Voice Conversion Model Based on Feature Disentanglement and Enhancement
Zihao Wang, Le Ma, Yongsheng Feng +3
Singing voice conversion (SVC) aims to convert a singer's voice to another singer's from a reference audio while keeping the original semantics. However, existing SVC methods can h…