44 citations · 45 across the 4 of their papers we have counts for
4 papers
A Recurrent Variational Autoencoder for Speech Enhancement
Simon Leglaive, Xavier Alameda-Pineda, Laurent Girin +1
This paper presents a generative approach to speech enhancement based on a recurrent variational autoencoder (RVAE). The deep generative speech model is trained using clean speech…
Extended Gaze Following: Detecting Objects in Videos Beyond the Camera Field of View
Benoit Massé, Stéphane Lathuilière, Pablo Mesejo +1
In this paper we address the problems of detecting objects of interest in a video and of estimating their locations, solely from the gaze directions of people present in the video.…
Speech enhancement with variational autoencoders and alpha-stable distributions
Simon Leglaive, Umut Simsekli, Antoine Liutkus +2
This paper focuses on single-channel semi-supervised speech enhancement. We learn a speaker-independent deep generative speech model using the framework of variational autoencoders…
A cascaded multiple-speaker localization and tracking system
Xiaofei Li, Yutong Ban, Laurent Girin +2
This paper presents an online multiple-speaker localization and tracking method, as the INRIA-Perception contribution to the LOCATA Challenge 2018. First, the recursive least-squar…