2 papers
cs.SD2022
iSTFTNet: Fast and Lightweight Mel-Spectrogram Vocoder Incorporating Inverse Short-Time Fourier Transform
Takuhiro Kaneko, Kou Tanaka, Hirokazu Kameoka +1
In recent text-to-speech synthesis and voice conversion systems, a mel-spectrogram is commonly applied as an intermediate representation, and the necessity for a mel-spectrogram vo…
stat.ML2018
Generalized Multichannel Variational Autoencoder for Underdetermined Source Separation
Shogo Seki, Hirokazu Kameoka, Li Li +2
This paper deals with a multichannel audio source separation problem under underdetermined conditions. Multichannel Non-negative Matrix Factorization (MNMF) is one of powerful appr…