23 citations · 60 across the 6 of their papers we have counts for
8 papers
ALIP: Adaptive Language-Image Pre-training with Synthetic Caption
Kaicheng Yang, Jiankang Deng, Xiang An +5
Contrastive Language-Image Pre-training (CLIP) has significantly boosted the performance of various vision-language tasks by scaling up the dataset with image-text pairs collected…
Unicom: Universal and Compact Representation Learning for Image Retrieval
Xiang An, Jiankang Deng, Kaicheng Yang +5
Modern image retrieval methods typically rely on fine-tuning pre-trained encoders to extract image-level descriptors. However, the most widely used models are pre-trained on ImageN…
Killing Two Birds with One Stone:Efficient and Robust Training of Face Recognition CNNs by Partial FC
Xiang An, Jiankang Deng, Jia Guo +4
Learning discriminative deep feature embeddings by using million-scale in-the-wild datasets and margin-based softmax loss is the current state-of-the-art approach for face recognit…
Point Cloud Registration using Representative Overlapping Points
Lifa Zhu, Dongrui Liu, Changwei Lin +4
3D point cloud registration is a fundamental task in robotics and computer vision. Recently, many learning-based point cloud registration methods based on correspondences have emer…
DeepText: A Unified Framework for Text Proposal Generation and Text Detection in Natural Images
Zhuoyao Zhong, Lianwen Jin, Shuye Zhang +1
In this paper, we develop a novel unified framework called DeepText for text region proposal generation and text detection in natural images via a fully convolutional neural networ…
Fully Convolutional Recurrent Network for Handwritten Chinese Text Recognition
Zecheng Xie, Zenghui Sun, Lianwen Jin +2
This paper proposes an end-to-end framework, namely fully convolutional recurrent network (FCRN) for handwritten Chinese text recognition (HCTR). Unlike traditional methods that re…