2 papers
cs.CV2025
AlignCAT: Visual-Linguistic Alignment of Category and Attribute for Weakly Supervised Visual Grounding
Yidan Wang, Chenyi Zhuang, Wutao Liu +2
Weakly supervised visual grounding (VG) aims to locate objects in images based on text descriptions. Despite significant progress, existing methods lack strong cross-modal reasonin…
cs.CV2025
First RAG, Second SEG: A Training-Free Paradigm for Camouflaged Object Detection
Wutao Liu, YiDan Wang, Pan Gao
Camouflaged object detection (COD) poses a significant challenge in computer vision due to the high similarity between objects and their backgrounds. Existing approaches often rely…