2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2025★ 2 cited
Multilingual Text-to-Image Person Retrieval via Bidirectional Relation Reasoning and Aligning
Min Cao, Xinyu Zhou, Ding Jiang +3
Text-to-image person retrieval (TIPR) aims to identify the target person using textual descriptions, facing challenge in modality heterogeneity. Prior works have attempted to addre…
cs.CR2025
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
Mang Ye, Xuankun Rong, Wenke Huang +3
With the rapid advancement of Large Vision-Language Models (LVLMs), ensuring their safety has emerged as a crucial area of research. This survey provides a comprehensive analysis o…
cs.CV2024
All in One Framework for Multimodal Re-identification in the Wild
He Li, Mang Ye, Ming Zhang +1
In Re-identification (ReID), recent advancements yield noteworthy progress in both unimodal and cross-modal retrieval tasks. However, the challenge persists in developing a unified…