3 papers
cs.CV2026
Learning Language-Driven Sequence-Level Modal-Invariant Representations for Video-Based Visible-Infrared Person Re-Identification
Xiaomei Yang, Antai Liu, Xizhan Gao +3
The core of video-based visible-infrared person re-identification (VVI-ReID) lies in learning sequence-level modal-invariant representations across different modalities. Recent res…
cs.CV2025★ 1 cited
CLIP4VI-ReID: Learning Modality-shared Representations via CLIP Semantic Bridge for Visible-Infrared Person Re-identification
Xiaomei Yang, Xizhan Gao, Sijie Niu +4
This paper proposes a novel CLIP-driven modality-shared representation learning network named CLIP4VI-ReID for VI-ReID task, which consists of Text Semantic Generation (TSG), Infra…
cs.CV2025
Normality Prior Guided Multi-Semantic Fusion Network for Unsupervised Image Anomaly Detection
Muhao Xu, Xueying Zhou, Xizhan Gao +3
Recently, detecting logical anomalies is becoming a more challenging task compared to detecting structural ones. Existing encoder decoder based methods typically compress inputs in…