134 citations · 335 across the 4 of their papers we have counts for
4 papers
Exploring Neural Transducers for End-to-End Speech Recognition
Eric Battenberg, Jitong Chen, Rewon Child +8
In this work, we perform an empirical comparison among the CTC, RNN-Transducer, and attention-based Seq2Seq models for end-to-end speech recognition. We show that, without any lang…
DeepID-Net: multi-stage and deformable deep convolutional neural networks for object detection
Wanli Ouyang, Ping Luo, Xingyu Zeng +12
In this paper, we propose multi-stage and deformable deep convolutional neural networks for object detection. This new deep learning object detection diagram has innovations in mul…
Deep Learning Multi-View Representation for Face Recognition
Zhenyao Zhu, Ping Luo, Xiaogang Wang +1
Various factors, such as identities, views (poses), and illuminations, are coupled in face images. Disentangling the identity and view representations is a major challenge in face…
Recover Canonical-View Faces in the Wild with Deep Neural Networks
Zhenyao Zhu, Ping Luo, Xiaogang Wang +1
Face images in the wild undergo large intra-personal variations, such as poses, illuminations, occlusions, and low resolutions, which cause great challenges to face-related applica…