collaborators

6 papers

cs.CV2025

Image Recognition with Online Lightweight Vision Transformer: A Survey

Zherui Zhang, Rongtao Xu, Jie Zhou +8

The Transformer architecture has achieved significant success in natural language processing, motivating its adaptation to computer vision tasks. Unlike convolutional neural networ…

cs.CV2025

SAMamba: Adaptive State Space Modeling with Hierarchical Vision for Infrared Small Target Detection

Wenhao Xu, Shuchen Zheng, Changwei Wang +4

Infrared small target detection (ISTD) is vital for long-range surveillance in military, maritime, and early warning applications. ISTD is challenged by targets occupying less than…

cs.CV2025

FDBPL: Faster Distillation-Based Prompt Learning for Region-Aware Vision-Language Models Adaptation

Zherui Zhang, Jiaxin Wu, Changwei Wang +6

Prompt learning as a parameter-efficient method that has been widely adopted to adapt Vision-Language Models (VLMs) to downstream tasks. While hard-prompt design requires domain ex…

cs.CV2025

CAE-DFKD: Bridging the Transferability Gap in Data-Free Knowledge Distillation

Zherui Zhang, Changwei Wang, Rongtao Xu +4

Data-Free Knowledge Distillation (DFKD) enables the knowledge transfer from the given pre-trained teacher network to the target student model without access to the real training da…

cs.CV2025

Focus on Local: Finding Reliable Discriminative Regions for Visual Place Recognition

Changwei Wang, Shunpeng Chen, Yukun Song +11

Visual Place Recognition (VPR) is aimed at predicting the location of a query image by referencing a database of geotagged images. For VPR task, often fewer discriminative local re…

cs.CV2024

Generalization Boosted Adapter for Open-Vocabulary Segmentation

Wenhao Xu, Changwei Wang, Xuxiang Feng +5

Vision-language models (VLMs) have demonstrated remarkable open-vocabulary object recognition capabilities, motivating their adaptation for dense prediction tasks like segmentation…