3 citations · 11 across the 11 of their papers we have counts for
15 papers
Speaking-Rate-Controllable HiFi-GAN Using Feature Interpolation
Detai Xin, Shinnosuke Takamichi, Takuma Okamoto +2
This paper presents a speaking-rate-controllable HiFi-GAN neural vocoder. Original HiFi-GAN is a high-fidelity, computationally efficient, and tiny-footprint neural vocoder. We att…
Partial Coupling of Optimal Transport for Spoken Language Identification
Xugang Lu, Peng Shen, Yu Tsao +1
In order to reduce domain discrepancy to improve the performance of cross-domain spoken language identification (SLID) system, as an unsupervised domain adaptation (UDA) method, we…
Siamese Neural Network with Joint Bayesian Model Structure for Speaker Verification
Xugang Lu, Peng Shen, Yu Tsao +1
Generative probability models are widely used for speaker verification (SV). However, the generative models are lack of discriminative feature selection ability. As a hypothesis te…
Predicting and Attending to Damaging Collisions for Placing Everyday Objects in Photo-Realistic Simulations
Aly Magassouba, Komei Sugiura, Angelica Nakayama +4
Placing objects is a fundamental task for domestic service robots (DSRs). Thus, inferring the collision-risk before a placing motion is crucial for achieving the requested task. Th…
Unsupervised neural adaptation model based on optimal transport for spoken language identification
Xugang Lu, Peng Shen, Yu Tsao +1
Due to the mismatch of statistical distributions of acoustic speech between training and testing sets, the performance of spoken language identification (SLID) could be drastically…
Alleviating the Burden of Labeling: Sentence Generation by Attention Branch Encoder-Decoder Network
Tadashi Ogura, Aly Magassouba, Komei Sugiura +4
Domestic service robots (DSRs) are a promising solution to the shortage of home care workers. However, one of the main limitations of DSRs is their inability to interact naturally…