30 citations · 103 across the 15 of their papers we have counts for
7 papers · 1 filter
Hierarchical Semantic Tree Concept Whitening for Interpretable Image Classification
Haixing Dai, Lu Zhang, Lin Zhao +9
With the popularity of deep neural networks (DNNs), model interpretability is becoming a critical concern. Many approaches have been developed to tackle the problem through post-ho…
Review of Large Vision Models and Visual Prompt Engineering
Jiaqi Wang, Zhengliang Liu, Lin Zhao +18
Visual prompt engineering is a fundamental technology in the field of visual and image Artificial General Intelligence, serving as a key component for achieving zero-shot capabilit…
SAMAug: Point Prompt Augmentation for Segment Anything Model
Haixing Dai, Chong Ma, Zhiling Yan +14
This paper introduces SAMAug, a novel visual point augmentation method for the Segment Anything Model (SAM) that enhances interactive image segmentation performance. SAMAug generat…
Instruction-ViT: Multi-Modal Prompts for Instruction Learning in ViT
Zhenxiang Xiao, Yuzhong Chen, Lu Zhang +14
Prompts have been proven to play a crucial role in large language models, and in recent years, vision models have also been using prompts to improve scalability for multiple downst…
BI AVAN: Brain inspired Adversarial Visual Attention Network
Heng Huang, Lin Zhao, Xintao Hu +4
Visual attention is a fundamental mechanism in the human brain, and it inspires the design of attention mechanisms in deep neural networks. However, most of the visual attention st…
Eye-gaze-guided Vision Transformer for Rectifying Shortcut Learning
Chong Ma, Lin Zhao, Yuzhong Chen +15
Learning harmful shortcuts such as spurious correlations and biases prevents deep neural networks from learning the meaningful and useful representations, thus jeopardizing the gen…