2 papers
cs.SD2023
Symbolic & Acoustic: Multi-domain Music Emotion Modeling for Instrumental Music
Kexin Zhu, Xulong Zhang, Jianzong Wang +2
Music Emotion Recognition involves the automatic identification of emotional elements within music tracks, and it has garnered significant attention due to its broad applicability…
cs.SD2022
Improving Speech Representation Learning via Speech-level and Phoneme-level Masking Approach
Xulong Zhang, Jianzong Wang, Ning Cheng +2
Recovering the masked speech frames is widely applied in speech representation learning. However, most of these models use random masking in the pre-training. In this work, we prop…