2 citations · 2 across the 2 of their papers we have counts for
4 papers
Polyphonic pitch detection with convolutional recurrent neural networks
Carl Thomé, Sven Ahlbäck
Recent directions in automatic speech recognition (ASR) research have shown that applying deep learning models from image recognition challenges in computer vision is beneficial. A…
Computational Pronunciation Analysis in Sung Utterances
Emir Demirel, Sven Ahlback, Simon Dixon
Recent automatic lyrics transcription (ALT) approaches focus on building stronger acoustic models or in-domain language models, while the pronunciation aspect is seldom touched upo…
Automatic Lyrics Transcription using Dilated Convolutional Neural Networks with Self-Attention
Emir Demirel, Sven Ahlback, Simon Dixon
Speech recognition is a well developed research field so that the current state of the art systems are being used in many applications in the software industry, yet as by today, th…
Investigating kernel shapes and skip connections for deep learning-based harmonic-percussive separation
Carlos Lordelo, Emmanouil Benetos, Simon Dixon +1
In this paper we propose an efficient deep learning encoder-decoder network for performing Harmonic-Percussive Source Separation (HPSS). It is shown that we are able to greatly red…