14 citations · 14 across the 2 of their papers we have counts for
2 papers
cs.SD2024
The Interpretation Gap in Text-to-Music Generation Models
Yongyi Zang, Yixiao Zhang
Large-scale text-to-music generation models have significantly enhanced music creation capabilities, offering unprecedented creative freedom. However, their ability to collaborate…
eess.AS2024★ 14 cited
CtrSVDD: A Benchmark Dataset and Baseline Analysis for Controlled Singing Voice Deepfake Detection
Yongyi Zang, Jiatong Shi, You Zhang +8
Recent singing voice synthesis and conversion advancements necessitate robust singing voice deepfake detection (SVDD) models. Current SVDD datasets face challenges due to limited c…