6 papers
Image Recognition with Online Lightweight Vision Transformer: A Survey
Zherui Zhang, Rongtao Xu, Jie Zhou +8
The Transformer architecture has achieved significant success in natural language processing, motivating its adaptation to computer vision tasks. Unlike convolutional neural networ…
SAMamba: Adaptive State Space Modeling with Hierarchical Vision for Infrared Small Target Detection
Wenhao Xu, Shuchen Zheng, Changwei Wang +4
Infrared small target detection (ISTD) is vital for long-range surveillance in military, maritime, and early warning applications. ISTD is challenged by targets occupying less than…
FDBPL: Faster Distillation-Based Prompt Learning for Region-Aware Vision-Language Models Adaptation
Zherui Zhang, Jiaxin Wu, Changwei Wang +6
Prompt learning as a parameter-efficient method that has been widely adopted to adapt Vision-Language Models (VLMs) to downstream tasks. While hard-prompt design requires domain ex…
CAE-DFKD: Bridging the Transferability Gap in Data-Free Knowledge Distillation
Zherui Zhang, Changwei Wang, Rongtao Xu +4
Data-Free Knowledge Distillation (DFKD) enables the knowledge transfer from the given pre-trained teacher network to the target student model without access to the real training da…
Focus on Local: Finding Reliable Discriminative Regions for Visual Place Recognition
Changwei Wang, Shunpeng Chen, Yukun Song +11
Visual Place Recognition (VPR) is aimed at predicting the location of a query image by referencing a database of geotagged images. For VPR task, often fewer discriminative local re…
Generalization Boosted Adapter for Open-Vocabulary Segmentation
Wenhao Xu, Changwei Wang, Xuxiang Feng +5
Vision-language models (VLMs) have demonstrated remarkable open-vocabulary object recognition capabilities, motivating their adaptation for dense prediction tasks like segmentation…