7 citations · 7 across the 2 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
Ye Bai, Chenxing Li, Hao Li +2
In short video and live broadcasts, speech, singing voice, and background music often overlap and obscure each other. This complexity creates difficulties in structuring and recogn…
cs.SD2021★ 7 cited
MELONS: generating melody with long-term structure using transformers and structure graph
Yi Zou, Pei Zou, Yi Zhao +3
The creation of long melody sequences requires effective expression of coherent musical structure. However, there is no clear representation of musical structure. Recent works on m…