1 paper
Wang Zhou, Boran Duan, Haojun Ai +2
Recent vision-language models such as CLIP provide strong cross-modal alignment, but current CLIP-guided ReID pipelines rely on global features and fixed prompts. This limits their…