3 papers
cs.CV2026
Rethinking Multi-Label Image Classification With Deep Learning: Taxonomy, Challenge, and Outlook
Xuelin Zhu, Xiu-Shen Wei, Jiawei Ge +2
Multi-label image classification (MLIC), a fundamental task in computer vision, focuses on identifying multiple objects or concepts within an image, underpinning numerous read-worl…
cs.CV2026
Debate-Enhanced Pseudo Labeling and Frequency-Aware Progressive Debiasing for Weakly-Supervised Camouflaged Object Detection with Scribble Annotations
Jiawei Ge, Jiuxin Cao, Xinyi Li +5
Weakly-Supervised Camouflaged Object Detection (WSCOD) aims to locate and segment objects that are visually concealed within their surrounding scenes, relying solely on sparse supe…
cs.CV2025
Denoise-then-Retrieve: Text-Conditioned Video Denoising for Video Moment Retrieval
Weijia Liu, Jiuxin Cao, Bo Miao +6
Current text-driven Video Moment Retrieval (VMR) methods encode all video clips, including irrelevant ones, disrupting multimodal alignment and hindering optimization. To this end,…