4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CL2019
Generating Synthetic Audio Data for Attention-Based Speech Recognition Systems
Nick Rossenbach, Albert Zeyer, Ralf Schlüter +1
Recent advances in text-to-speech (TTS) led to the development of flexible multi-speaker end-to-end TTS systems. We extend state-of-the-art attention-based automatic speech recogni…
cs.CL2019★ 4 cited
Learning Bilingual Sentence Embeddings via Autoencoding and Computing Similarities with a Multilayer Perceptron
Yunsu Kim, Hendrik Rosendahl, Nick Rossenbach +3
We propose a novel model architecture and training algorithm to learn bilingual sentence embeddings from a combination of parallel and monolingual data. Our method connects autoenc…