2 papers
cs.SD2021
AvaTr: One-Shot Speaker Extraction with Transformers
Shell Xu Hu, Md Rifat Arefin, Viet-Nhat Nguyen +3
To extract the voice of a target speaker when mixed with a variety of other sounds, such as white and ambient noises or the voices of interfering speakers, we extend the Transforme…
cs.CL2019
Instance-Based Model Adaptation For Direct Speech Translation
Mattia Antonino Di Gangi, Viet-Nhat Nguyen, Matteo Negri +1
Despite recent technology advancements, the effectiveness of neural approaches to end-to-end speech-to-text translation is still limited by the paucity of publicly available traini…