46 citations · 46 across the 1 of their papers we have counts for
3 papers
Multichannel End-to-end Speech Recognition
Tsubasa Ochiai, Shinji Watanabe, Takaaki Hori +1
The field of speech recognition is in the midst of a paradigm shift: end-to-end neural networks are challenging the dominance of hidden Markov models as a core technology. Using an…
Single-Channel Multi-Speaker Separation using Deep Clustering
Yusuf Isik, Jonathan Le Roux, Zhuo Chen +2
Deep clustering is a recently introduced deep learning architecture that uses discriminatively trained embeddings as the basis for clustering. It was recently applied to spectrogra…
Global-Local Face Upsampling Network
Oncel Tuzel, Yuichi Taguchi, John R. Hershey
Face hallucination, which is the task of generating a high-resolution face image from a low-resolution input image, is a well-studied problem that is useful in widespread applicati…