3 citations · 3 across the 3 of their papers we have counts for
11 papers
Room Impulse Responses help attackers to evade Deep Fake Detection
Hieu-Thi Luong, Duc-Tuan Truong, Kong Aik Lee +1
The ASVspoof 2021 benchmark, a widely-used evaluation framework for anti-spoofing, consists of two subsets: Logical Access (LA) and Deepfake (DF), featuring samples with varied cod…
Preliminary study on using vector quantization latent spaces for TTS/VC systems with consistent performance
Hieu-Thi Luong, Junichi Yamagishi
Generally speaking, the main objective when training a neural speech synthesis system is to synthesize natural and expressive speech from the output layer of the neural network wit…
Latent linguistic embedding for cross-lingual text-to-speech and voice conversion
Hieu-Thi Luong, Junichi Yamagishi
As the recently proposed voice cloning system, NAUTILUS, is capable of cloning unseen voices using untranscribed speech, we investigate the feasibility of using it to develop a uni…
NAUTILUS: a Versatile Voice Cloning System
Hieu-Thi Luong, Junichi Yamagishi
We introduce a novel speech synthesis system, called NAUTILUS, that can generate speech with a target voice either from a text input or a reference utterance of an arbitrary source…
Bootstrapping non-parallel voice conversion from speaker-adaptive text-to-speech
Hieu-Thi Luong, Junichi Yamagishi
Voice conversion (VC) and text-to-speech (TTS) are two tasks that share a similar objective, generating speech with a target voice. However, they are usually developed independentl…
A Unified Speaker Adaptation Method for Speech Synthesis using Transcribed and Untranscribed Speech with Backpropagation
Hieu-Thi Luong, Junichi Yamagishi
By representing speaker characteristic as a single fixed-length vector extracted solely from speech, we can train a neural multi-speaker speech synthesis model by conditioning the…