2 citations · 2 across the 3 of their papers we have counts for
4 papers
MAP-Music2Vec: A Simple and Effective Baseline for Self-Supervised Music Audio Representation Learning
Yizhi Li, Ruibin Yuan, Ge Zhang +11
The deep learning community has witnessed an exponentially growing interest in self-supervised learning (SSL). However, it still remains unexplored how to build a framework for lea…
Learnable Front Ends Based on Temporal Modulation for Music Tagging
Yinghao Ma, Richard M. Stern
While end-to-end systems are becoming popular in auditory signal processing including automatic music tagging, models using raw audio as input needs a large amount of data and comp…
PL-EESR: Perceptual Loss Based END-TO-END Robust Speaker Representation Extraction
Yi Ma, Kong Aik Lee, Ville Hautamaki +1
Speech enhancement aims to improve the perceptual quality of the speech signal by suppression of the background noise. However, excessive suppression may lead to speech distortion…
A Transformer Based Pitch Sequence Autoencoder with MIDI Augmentation
Mingshuo Ding, Yinghao Ma
Despite recent achievements of deep learning automatic music generation algorithms, few approaches have been proposed to evaluate whether a single-track music excerpt is composed b…