154 citations · 848 across the 33 of their papers we have counts for
14 papers · 1 filter
Vision-Infused Deep Audio Inpainting
Hang Zhou, Ziwei Liu, Xudong Xu +2
Multi-modality perception is essential to develop interactive intelligence. In this work, we consider a new task of visual information-infused audio inpainting, \ie synthesizing mi…
TextSR: Content-Aware Text Super-Resolution Guided by Recognition
Wenjia Wang, Enze Xie, Peize Sun +4
Scene text recognition has witnessed rapid development with the advance of convolutional neural networks. Nonetheless, most of the previous methods may not work well in recognizing…
PolarMask: Single Shot Instance Segmentation with Polar Representation
Enze Xie, Peize Sun, Xiaoge Song +4
In this paper, we introduce an anchor-box free and single shot instance segmentation method, which is conceptually simple, fully convolutional and can be used as a mask prediction…
Scale Calibrated Training: Improving Generalization of Deep Networks via Scale-Specific Normalization
Zhuoran Yu, Aojun Zhou, Yukun Ma +3
Standard convolutional neural networks(CNNs) require consistent image resolutions in both training and testing phase. However, in practice, testing with smaller image sizes is nece…
Fashion Retrieval via Graph Reasoning Networks on a Similarity Pyramid
Zhanghui Kuang, Yiming Gao, Guanbin Li +4
Matching clothing images from customers and online shopping stores has rich applications in E-commerce. Existing algorithms encoded an image as a global feature vector and performe…
Deep Self-Learning From Noisy Labels
Jiangfan Han, Ping Luo, Xiaogang Wang
ConvNets achieve good results when training from clean data, but learning from noisy labels significantly degrades performances and remains challenging. Unlike previous works const…