43 citations · 145 across the 9 of their papers we have counts for
8 papers
HiVLP: Hierarchical Vision-Language Pre-Training for Fast Image-Text Retrieval
Feilong Chen, Xiuyi Chen, Jiaxin Shi +3
In the past few years, the emergence of vision-language pre-training (VLP) has brought cross-modal retrieval to a new era. However, due to the latency and computation demand, it is…
Differentiable Convolution Search for Point Cloud Processing
Xing Nie, Yongcheng Liu, Shaohong Chen +6
Exploiting convolutional neural networks for point cloud processing is quite challenging, due to the inherent irregular distribution and discrete shape representation of point clou…
What Is Considered Complete for Visual Recognition?
Lingxi Xie, Xiaopeng Zhang, Longhui Wei +2
This is an opinion paper. We hope to deliver a key message that current visual recognition systems are far from complete, i.e., recognizing everything that human can recognize, yet…
Can Semantic Labels Assist Self-Supervised Visual Representation Learning?
Longhui Wei, Lingxi Xie, Jianzhong He +5
Recently, contrastive learning has largely advanced the progress of unsupervised visual representation learning. Pre-trained on ImageNet, some self-supervised algorithms reported h…
Weight-Sharing Neural Architecture Search: A Battle to Shrink the Optimization Gap
Lingxi Xie, Xin Chen, Kaifeng Bi +8
Neural architecture search (NAS) has attracted increasing attentions in both academia and industry. In the early age, researchers mostly applied individual search methods which sam…
Cross-Modality Paired-Images Generation for RGB-Infrared Person Re-Identification
Guan-An Wang, Tianzhu Zhang. Yang Yang, Jian Cheng +3
RGB-Infrared (IR) person re-identification is very challenging due to the large cross-modality variations between RGB and IR images. The key solution is to learn aligned features t…