434 citations
- Alibaba Group (China)CN14 papers
- Alibaba Group (Cayman Islands)KY6 papers
- Hong Kong Polytechnic UniversityHK5 papers
- Nanyang Technological UniversitySG4 papers
- Zhejiang UniversityCN4 papers
- Chinese Academy of SciencesCN2 papers
- Chinese University of Hong KongHK2 papers
- Harbin Institute of TechnologyCN2 papers
- Peng Cheng LaboratoryCN2 papers
- Renmin University of ChinaCN2 papers
- Tencent (China)CN2 papers
- Alibaba Group (United States)US1 paper
22 papers
Region-adaptive Texture Enhancement for Detailed Person Image Synthesis
Lingbo Yang, Pan Wang, Xinfeng Zhang +6
The ability to produce convincing textural details is essential for the fidelity of synthesized person images. However, existing methods typically follow a ``warping-based'' strate…
Predict-then-Decide: A Predictive Approach for Wait or Answer Task in Dialogue Systems
Zehao Lin, Shaobo Cui, Guodun Li +6
Different people have different habits of describing their intents in conversations. Some people tend to deliberate their intents in several successive utterances, i.e., they use s…
Simplified Self-Attention for Transformer-based End-to-End Speech Recognition
Haoneng Luo, Shiliang Zhang, Ming Lei +1
Transformer models have been introduced into end-to-end speech recognition with state-of-the-art performance on various tasks owing to their superiority in modeling long-term depen…
Learning in the Frequency Domain
Kai Xu, Minghai Qin, Fei Sun +3
Deep neural networks have achieved remarkable success in computer vision tasks. Existing neural networks mainly operate in the spatial domain with fixed input sizes. For practical…
Cross-modality Person re-identification with Shared-Specific Feature Transfer
Yan Lu, Yue Wu, Bin Liu +4
Cross-modality person re-identification (cm-ReID) is a challenging but key technology for intelligent video analysis. Existing works mainly focus on learning common representation…
Visual Commonsense R-CNN
Tan Wang, Jianqiang Huang, Hanwang Zhang +1
We present a novel unsupervised feature representation learning method, Visual Commonsense Region-based Convolutional Neural Network (VC R-CNN), to serve as an improved visual regi…