2 papers
cs.CV2024
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
Xiao Wang, Jiandong Jin, Chenglong Li +3
Existing pedestrian attribute recognition (PAR) algorithms adopt pre-trained CNN (e.g., ResNet) as their backbone network for visual feature learning, which might obtain sub-optima…
cs.CV2024
Learning Adaptive Fusion Bank for Multi-modal Salient Object Detection
Kunpeng Wang, Zhengzheng Tu, Chenglong Li +2
Multi-modal salient object detection (MSOD) aims to boost saliency detection performance by integrating visible sources with depth or thermal infrared ones. Existing methods genera…