activity
20172024
most citedJSUT corpus: free large-scale Japanese speech corpus for end-to-end speech synthesis

88 citations · 197 across the 76 of their papers we have counts for

collaborators
Showing 2021 · cs.SDShow all

10 papers · 2 filters

cs.SD2021

Mean-square-error-based secondary source placement in sound field synthesis with prior information on desired field

Keisuke Kimura, Shoichi Koyama, Natsuki Ueno +1

A method of optimizing secondary source placement in sound field synthesis is proposed. Such an optimization method will be useful when the allowable placement region and available…

cs.SD2021★ 2 cited

Kernel Learning For Sound Field Estimation With L1 and L2 Regularizations

Ryosuke Horiuchi, Shoichi Koyama, Juliano G. C. Ribeiro +2

A method to estimate an acoustic field from discrete microphone measurements is proposed. A kernel-interpolation-based method using the kernel function formulated for sound field i…

cs.SD2021

Low-Latency Incremental Text-to-Speech Synthesis with Distilled Context Prediction Network

Takaaki Saeki, Shinnosuke Takamichi, Hiroshi Saruwatari

Incremental text-to-speech (TTS) synthesis generates utterances in small linguistic units for the sake of real-time and low-latency applications. We previously proposed an incremen…

cs.SD2021

Speech Enhancement by Noise Self-Supervised Rank-Constrained Spatial Covariance Matrix Estimation via Independent Deeply Learned Matrix Analysis

Sota Misawa, Norihiro Takamune, Tomohiko Nakamura +4

Rank-constrained spatial covariance matrix estimation (RCSCME) is a method for the situation that the directional target speech and the diffuse noise are mixed. In conventional RCS…

cs.SD2021

Multichannel Audio Source Separation with Independent Deeply Learned Matrix Analysis Using Product of Source Models

Takuya Hasumi, Tomohiko Nakamura, Norihiro Takamune +4

Independent deeply learned matrix analysis (IDLMA) is one of the state-of-the-art multichannel audio source separation methods using the source power estimation based on deep neura…

cs.SD2021

Prior Distribution Design for Music Bleeding-Sound Reduction Based on Nonnegative Matrix Factorization

Yusaku Mizobuchi, Daichi Kitamura, Tomohiko Nakamura +3

When we place microphones close to a sound source near other sources in audio recording, the obtained audio signal includes undesired sound from the other sources, which is often c…