69 citations · 83 across the 9 of their papers we have counts for
16 papers
Generalized Fast Multichannel Nonnegative Matrix Factorization Based on Gaussian Scale Mixtures for Blind Source Separation
Mathieu Fontaine, Kouhei Sekiguchi, Aditya Nugraha +2
This paper describes heavy-tailed extensions of a state-of-the-art versatile blind source separation method called fast multichannel nonnegative matrix factorization (FastMNMF) fro…
Global Structure-Aware Drum Transcription Based on Self-Attention Mechanisms
Ryoto Ishizuka, Ryo Nishikimi, Kazuyoshi Yoshii
This paper describes an automatic drum transcription (ADT) method that directly estimates a tatum-level drum score from a music signal, in contrast to most conventional ADT methods…
Tatum-Level Drum Transcription Based on a Convolutional Recurrent Neural Network with Language Model-Based Regularized Training
Ryoto Ishizuka, Ryo Nishikimi, Eita Nakamura +1
This paper describes a neural drum transcription method that detects from music signals the onset times of drums at the level, where tatum times are assumed to be…
The MIDI Degradation Toolkit: Symbolic Music Augmentation and Correction
Andrew McLeod, James Owers, Kazuyoshi Yoshii
In this paper, we introduce the MIDI Degradation Toolkit (MDTK), containing functions which take as input a musical excerpt (a set of notes with pitch, onset time, and duration), a…
End-to-end Music-mixed Speech Recognition
Jeongwoo Woo, Masato Mimura, Kazuyoshi Yoshii +1
Automatic speech recognition (ASR) in multimedia content is one of the promising applications, but speech data in this kind of content are frequently mixed with background music, w…
Non-Local Musical Statistics as Guides for Audio-to-Score Piano Transcription
Kentaro Shibata, Eita Nakamura, Kazuyoshi Yoshii
We present an automatic piano transcription system that converts polyphonic audio recordings into musical scores. This has been a long-standing problem of music information process…