activity
20172022
most citedUnsupervised Speech Enhancement Based on Multichannel NMF-Informed Beamforming for Noise-Robust Automatic Speech Recognition

69 citations · 83 across the 9 of their papers we have counts for

collaborators

16 papers

cs.SD20229 cited

Generalized Fast Multichannel Nonnegative Matrix Factorization Based on Gaussian Scale Mixtures for Blind Source Separation

Mathieu Fontaine, Kouhei Sekiguchi, Aditya Nugraha +2

This paper describes heavy-tailed extensions of a state-of-the-art versatile blind source separation method called fast multichannel nonnegative matrix factorization (FastMNMF) fro…

cs.SD2021

Global Structure-Aware Drum Transcription Based on Self-Attention Mechanisms

Ryoto Ishizuka, Ryo Nishikimi, Kazuyoshi Yoshii

This paper describes an automatic drum transcription (ADT) method that directly estimates a tatum-level drum score from a music signal, in contrast to most conventional ADT methods…

cs.SD2020

Tatum-Level Drum Transcription Based on a Convolutional Recurrent Neural Network with Language Model-Based Regularized Training

Ryoto Ishizuka, Ryo Nishikimi, Eita Nakamura +1

This paper describes a neural drum transcription method that detects from music signals the onset times of drums at the level, where tatum times are assumed to be…

cs.SD20201 cited

The MIDI Degradation Toolkit: Symbolic Music Augmentation and Correction

Andrew McLeod, James Owers, Kazuyoshi Yoshii

In this paper, we introduce the MIDI Degradation Toolkit (MDTK), containing functions which take as input a musical excerpt (a set of notes with pitch, onset time, and duration), a…

eess.AS2020

End-to-end Music-mixed Speech Recognition

Jeongwoo Woo, Masato Mimura, Kazuyoshi Yoshii +1

Automatic speech recognition (ASR) in multimedia content is one of the promising applications, but speech data in this kind of content are frequently mixed with background music, w…

cs.SD2020

Non-Local Musical Statistics as Guides for Audio-to-Score Piano Transcription

Kentaro Shibata, Eita Nakamura, Kazuyoshi Yoshii

We present an automatic piano transcription system that converts polyphonic audio recordings into musical scores. This has been a long-standing problem of music information process…