12 citations · 19 across the 14 of their papers we have counts for
4 papers · 1 filter
Optimal Condition Training for Target Source Separation
Efthymios Tzinis, Gordon Wichern, Paris Smaragdis +1
Recent research has shown remarkable performance in leveraging multiple extraneous conditional and non-mutually exclusive semantic concepts for sound source separation, allowing th…
Extended Graph Temporal Classification for Multi-Speaker End-to-End ASR
Xuankai Chang, Niko Moritz, Takaaki Hori +2
Graph-based temporal classification (GTC), a generalized form of the connectionist temporal classification loss, was recently proposed to improve automatic speech recognition (ASR)…
Leveraging Low-Distortion Target Estimates for Improved Speech Enhancement
Zhong-Qiu Wang, Gordon Wichern, Jonathan Le Roux
A promising approach for multi-microphone speech separation involves two deep neural networks (DNN), where the predicted target speech from the first DNN is used to compute signal…
Convolutive Prediction for Reverberant Speech Separation
Zhong-Qiu Wang, Gordon Wichern, Jonathan Le Roux
We investigate the effectiveness of convolutive prediction, a novel formulation of linear prediction for speech dereverberation, for speaker separation in reverberant conditions. T…