activity
20222024
most citedGraphAdapter: Tuning Vision-Language Models With Dual Knowledge Graph

21 citations · 43 across the 8 of their papers we have counts for

collaborators

8 papers

cs.CV2024

Parameter-Efficient and Memory-Efficient Tuning for Vision Transformer: A Disentangled Approach

Taolin Zhang, Jiawang Bai, Zhihe Lu +4

Recent works on parameter-efficient transfer learning (PETL) show the potential to adapt a pre-trained Vision Transformer to downstream recognition tasks with only a few learnable…

cs.CV202321 cited

GraphAdapter: Tuning Vision-Language Models With Dual Knowledge Graph

Xin Li, Dongze Lian, Zhihe Lu +3

Adapter-style efficient transfer learning (ETL) has shown excellent performance in the tuning of vision-language models (VLMs) under the low-data regime, where only a few additiona…

cs.CV20236 cited

A Dive into SAM Prior in Image Restoration

Zeyu Xiao, Jiawang Bai, Zhihe Lu +1

The goal of image restoration (IR), a fundamental issue in computer vision, is to restore a high-quality (HQ) image from its degraded low-quality (LQ) observation. Multiple HQ solu…

cs.CV202314 cited

Can SAM Boost Video Super-Resolution?

Zhihe Lu, Zeyu Xiao, Jiawang Bai +2

The primary challenge in video super-resolution (VSR) is to handle large motions in the input frames, which makes it difficult to accurately aggregate information from multiple fra…

cs.CV2023

Reliable and Efficient Evaluation of Adversarial Robustness for Deep Hashing-Based Retrieval

Xunguang Wang, Jiawang Bai, Xinyue Xu +1

Deep hashing has been extensively applied to massive image retrieval due to its efficiency and effectiveness. Recently, several adversarial attacks have been presented to reveal th…

cs.CV20221 cited

Imperceptible and Robust Backdoor Attack in 3D Point Cloud

Kuofeng Gao, Jiawang Bai, Baoyuan Wu +2

With the thriving of deep learning in processing point cloud data, recent works show that backdoor attacks pose a severe security threat to 3D vision applications. The attacker inj…