activity
20192024
most citedDirect speech-to-speech translation with a sequence-to-sequence model

22 citations · 22 across the 5 of their papers we have counts for

collaborators

6 papers

eess.AS2024

Zero-shot Cross-lingual Voice Transfer for TTS

Fadi Biadsy, Youzheng Chen, Isaac Elias +4

In this paper, we introduce a zero-shot Voice Transfer (VT) module that can be seamlessly integrated into a multi-lingual Text-to-speech (TTS) system to transfer an individual's vo…

eess.AS2024

Hierarchical Recurrent Adapters for Efficient Multi-Task Adaptation of Large Speech Models

Tsendsuren Munkhdalai, Youzheng Chen, Khe Chai Sim +3

Parameter efficient adaptation methods have become a key mechanism to train large pre-trained models for downstream tasks. However, their per-task parameter overhead is considered…

cs.SD2022

Non-Parallel Voice Conversion for ASR Augmentation

Gary Wang, Andrew Rosenberg, Bhuvana Ramabhadran +4

Automatic speech recognition (ASR) needs to be robust to speaker differences. Voice Conversion (VC) modifies speaker characteristics of input speech. This is an attractive feature…

cs.CL2021

Residual Adapters for Parameter-Efficient ASR Adaptation to Atypical and Accented Speech

Katrin Tomanek, Vicky Zayats, Dirk Padfield +2

Automatic Speech Recognition (ASR) systems are often optimized to work best for speakers with canonical speech patterns. Unfortunately, these systems perform poorly when tested on…

cs.CL201922 cited

Direct speech-to-speech translation with a sequence-to-sequence model

Ye Jia, Ron J. Weiss, Fadi Biadsy +4

We present an attention-based sequence-to-sequence neural network which can directly translate speech from one language into speech in another language, without relying on an inter…

eess.AS2019

Parrotron: An End-to-End Speech-to-Speech Conversion Model and its Applications to Hearing-Impaired Speech and Speech Separation

Fadi Biadsy, Ron J. Weiss, Pedro J. Moreno +2

We describe Parrotron, an end-to-end-trained speech-to-speech conversion model that maps an input spectrogram directly to another spectrogram, without utilizing any intermediate di…