Publications (4)
CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
Shinji Watanabe, Michael Mandel, Jon Barker +18
Following the success of the 1st, 2nd, 3rd, 4th and 5th CHiME challenges we organize the 6th CHiME Speech Separation and Recognition Challenge (CHiME-6). The new challenge revisits…
Enhancement of Spatial Clustering-Based Time-Frequency Masks using LSTM Neural Networks
Felix Grezes, Zhaoheng Ni, Viet Anh Trinh +1
Recent works have shown that Deep Recurrent Neural Networks using the LSTM architecture can achieve strong single-channel speech enhancement by estimating time-frequency masks. How…
Combining Spatial Clustering with LSTM Speech Models for Multichannel Speech Enhancement
Felix Grezes, Zhaoheng Ni, Viet Anh Trinh +1
Recurrent neural networks using the LSTM architecture can achieve significant single-channel noise reduction. It is not obvious, however, how to apply them to multi-channel inputs…
Autotagging music with conditional restricted Boltzmann machines
Michael Mandel, Razvan Pascanu, Hugo Larochelle +1
This paper describes two applications of conditional restricted Boltzmann machines (CRBMs) to the task of autotagging music. The first consists of training a CRBM to predict tags t…