2 papers
cs.SD2023
Deep Attention-Based Alignment Network for Melody Generation from Incomplete Lyrics
Gurunath Reddy M, Zhe Zhang, Yi Yu +3
We propose a deep attention-based alignment network, which aims to automatically predict lyrics and melody with given incomplete lyrics as input in a way similar to the music creat…
cs.IR2021
Variational Autoencoder with CCA for Audio-Visual Cross-Modal Retrieval
Jiwei Zhang, Yi Yu, Suhua Tang +2
Cross-modal retrieval is to utilize one modality as a query to retrieve data from another modality, which has become a popular topic in information retrieval, machine learning, and…