1 paper
Tianlu Zheng, Yifan Zhang, Xiang An +3
Although Contrastive Language-Image Pre-training (CLIP) exhibits strong performance across diverse vision tasks, its application to person representation learning faces two critica…