5 papers
Cross-modal Proxy Evolving for OOD Detection with Vision-Language Models
Hao Tang, Yu Liu, Shuanglin Yan +3
Reliable zero-shot detection of out-of-distribution (OOD) inputs is critical for deploying vision-language models in open-world settings. However, the lack of labeled negatives in…
Diverse Semantics-Guided Feature Alignment and Decoupling for Visible-Infrared Person Re-Identification
Neng Dong, Shuanglin Yan, Liyan Zhang +1
Visible-Infrared Person Re-Identification (VI-ReID) is a challenging task due to the large modality discrepancy between visible and infrared images, which complicates the alignment…
ShapeSpeak: Body Shape-Aware Textual Alignment for Visible-Infrared Person Re-Identification
Shuanglin Yan, Neng Dong, Shuang Li +3
Visible-Infrared Person Re-identification (VIReID) aims to match visible and infrared pedestrian images, but the modality differences and the complexity of identity features make i…
Embedding and Enriching Explicit Semantics for Visible-Infrared Person Re-Identification
Neng Dong, Shuanglin Yan, Liyan Zhang +1
Visible-infrared person re-identification (VIReID) retrieves pedestrian images with the same identity across different modalities. Existing methods learn visual content solely from…
Prototypical Prompting for Text-to-image Person Re-identification
Shuanglin Yan, Jun Liu, Neng Dong +2
In this paper, we study the problem of Text-to-Image Person Re-identification (TIReID), which aims to find images of the same identity described by a text sentence from a pool of c…