most citedCross-speaker style transfer for text-to-speech using data augmentation

2 citations · 4 across the 4 of their papers we have counts for

collaborators

5 papers

math.NA20222 cited

Tensor-product space-time goal-oriented error control and adaptivity with partition-of-unity dual-weighted residuals for nonstationary flow problems

Julian Roth, Jan Philipp Thiele, Uwe Köcher +1

In this work, the dual-weighted residual method is applied to a space-time formulation of nonstationary Stokes and Navier-Stokes flow. Tensor-product space-time finite elements are…

eess.AS2022

Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module

Adam Gabryś, Goeric Huybrechts, Manuel Sam Ribeiro +6

State-of-the-art text-to-speech (TTS) systems require several hours of recorded speech data to generate high-quality synthetic speech. When using reduced amounts of training data,…

eess.AS20222 cited

Cross-speaker style transfer for text-to-speech using data augmentation

Manuel Sam Ribeiro, Julian Roth, Giulia Comini +3

We address the problem of cross-speaker style transfer for text-to-speech (TTS) using data augmentation via voice conversion. We assume to have a corpus of neutral non-expressive d…

cs.SD2021

Improving multi-speaker TTS prosody variance with a residual encoder and normalizing flows

Iván Vallés-Pérez, Julian Roth, Grzegorz Beringer +2

Text-to-speech systems recently achieved almost indistinguishable quality from human speech. However, the prosody of those systems is generally flatter than natural speech, produci…

math.NA2021

Neural network guided adjoint computations in dual weighted residual error estimation

Julian Roth, Max Schröder, Thomas Wick

In this work, we are concerned with neural network guided goal-oriented a posteriori error estimation and adaptivity using the dual weighted residual method. The primal problem is…