14 citations · 60 across the 17 of their papers we have counts for
20 papers
Sampling-Frequency-Independent Audio Source Separation Using Convolution Layer Based on Impulse Invariant Method
Koichi Saito, Tomohiko Nakamura, Kohei Yatabe +2
Audio source separation is often used as preprocessing of various applications, and one of its ultimate goals is to construct a single versatile model capable of dealing with the v…
Sparse time-frequency representation via atomic norm minimization
Tsubasa Kusano, Kohei Yatabe, Yasuhiro Oikawa
Nonstationary signals are commonly analyzed and processed in the time-frequency (T-F) domain that is obtained by the discrete Gabor transform (DGT). The T-F representation obtained…
Mixture of orthogonal sequences made from extended time-stretched pulses enables measurement of involuntary voice fundamental frequency response to pitch perturbation
Hideki Kawahara, Toshie Matsui, Kohei Yatabe +4
Auditory feedback plays an essential role in the regulation of the fundamental frequency of voiced sounds. The fundamental frequency also responds to auditory stimulation other tha…
Noisy-target Training: A Training Strategy for DNN-based Speech Enhancement without Clean Speech
Takuya Fujimura, Yuma Koizumi, Kohei Yatabe +1
Deep neural network (DNN)-based speech enhancement ordinarily requires clean speech signals as the training target. However, collecting clean signals is very costly because they mu…
Cascaded all-pass filters with randomized center frequencies and phase polarity for acoustic and speech measurement and data augmentation
Hideki Kawahara, Kohei Yatabe
We introduce a new member of TSP (Time Stretched Pulse) for acoustic and speech measurement infrastructure, based on a simple all-pass filter and systematic randomization. This new…
Self-supervised Neural Audio-Visual Sound Source Localization via Probabilistic Spatial Modeling
Yoshiki Masuyama, Yoshiaki Bando, Kohei Yatabe +3
Detecting sound source objects within visual observation is important for autonomous robots to comprehend surrounding environments. Since sounding objects have a large variety with…