10 citations · 10 across the 4 of their papers we have counts for
4 papers
Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation
Jarod Duret, Yannick Estève, Titouan Parcollet
Recent advancements in textless speech-to-speech translation systems have been driven by the adoption of self-supervised learning techniques. Although most state-of-the-art systems…
How Should We Extract Discrete Audio Tokens from Self-Supervised Models?
Pooneh Mousavi, Jarod Duret, Salah Zaiem +4
Discrete audio tokens have recently gained attention for their potential to bridge the gap between audio and language processing. Ideal audio tokens must preserve content, paraling…
Enhancing expressivity transfer in textless speech-to-speech translation
Jarod Duret, Benjamin O'Brien, Yannick Estève +1
Textless speech-to-speech translation systems are rapidly advancing, thanks to the integration of self-supervised learning techniques. However, existing state-of-the-art systems fa…
Direct Text to Speech Translation System using Acoustic Units
Victoria Mingote, Pablo Gimeno, Luis Vicente +3
This paper proposes a direct text to speech translation system using discrete acoustic units. This framework employs text in different source languages as input to generate speech…