4 papers · 1 filter
DecomCAM: Advancing Beyond Saliency Maps through Decomposition and Integration
Yuguang Yang, Runtang Guo, Sheng Wu +6
Interpreting complex deep networks, notably pre-trained vision-language models (VLMs), is a formidable challenge. Current Class Activation Map (CAM) methods highlight regions revea…
Self-Enhancement Improves Text-Image Retrieval in Foundation Visual-Language Models
Yuguang Yang, Yiming Wang, Shupeng Geng +4
The emergence of cross-modal foundation models has introduced numerous approaches grounded in text-image retrieval. However, on some domain-specific retrieval tasks, these models f…
Decom--CAM: Tell Me What You See, In Details! Feature-Level Interpretation via Decomposition Class Activation Map
Yuguang Yang, Runtang Guo, Sheng Wu +4
Interpretation of deep learning remains a very challenging problem. Although the Class Activation Map (CAM) is widely used to interpret deep model predictions by highlighting objec…
Generative Modeling in Structural-Hankel Domain for Color Image Inpainting
Zihao Li, Chunhua Wu, Shenglin Wu +3
In recent years, some researchers focused on using a single image to obtain a large number of samples through multi-scale features. This study intends to a brand-new idea that requ…