4 papers
A Toolkit for Joint Speaker Diarization and Identification with Application to Speaker-Attributed ASR
Giovanni Morrone, Enrico Zovato, Fabio Brugnara +2
We present a modular toolkit to perform joint speaker diarization and speaker identification. The toolkit can leverage on multiple models and algorithms which are defined in a conf…
Audio-Visual Target Speaker Enhancement on Multi-Talker Environment using Event-Driven Cameras
Ander Arriandiaga, Giovanni Morrone, Luca Pasa +2
We propose a method to address audio-visual target speaker enhancement in multi-talker environments using event-driven cameras. State of the art audio-visual speech separation meth…
An Analysis of Speech Enhancement and Recognition Losses in Limited Resources Multi-talker Single Channel Audio-Visual ASR
Luca Pasa, Giovanni Morrone, Leonardo Badino
In this paper, we analyzed how audio-visual speech enhancement can help to perform the ASR task in a cocktail party scenario. Therefore we considered two simple end-to-end LSTM-bas…
Face Landmark-based Speaker-Independent Audio-Visual Speech Enhancement in Multi-Talker Environments
Giovanni Morrone, Luca Pasa, Vadim Tikhanoff +3
In this paper, we address the problem of enhancing the speech of a speaker of interest in a cocktail party scenario when visual information of the speaker of interest is available.…