21 citations · 21 across the 3 of their papers we have counts for
3 papers
cs.CV2026
Acoustically Grounded Cost Learning for Open-Vocabulary Audio-Visual Semantic Segmentation
Tianrui Hui, Shaofei Huang, Qisong Han +6
Open-Vocabulary Audio-Visual Semantic Segmentation (OV-AVSS) aims to perform pixel-level segmentation of sound-emitting objects from an open set of categories. The previous method…
cs.CV2025
CAMeL: Cross-modality Adaptive Meta-Learning for Text-based Person Retrieval
Hang Yu, Jiahao Wen, Zhedong Zheng
Text-based person retrieval aims to identify specific individuals within an image database using textual descriptions. Due to the high cost of annotation and privacy protection, re…
cs.CV2024★ 21 cited
From Data Deluge to Data Curation: A Filtering-WoRA Paradigm for Efficient Text-based Person Search
Jintao Sun, Hao Fei, Zhedong Zheng +1
In text-based person search endeavors, data generation has emerged as a prevailing practice, addressing concerns over privacy preservation and the arduous task of manual annotation…