activity
20222026
most citedRadiology-Llama2: Best-in-Class Large Language Model for Radiology

30 citations · 76 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2023

Hierarchical Semantic Tree Concept Whitening for Interpretable Image Classification

Haixing Dai, Lu Zhang, Lin Zhao +9

With the popularity of deep neural networks (DNNs), model interpretability is becoming a critical concern. Many approaches have been developed to tackle the problem through post-ho…

cs.CV2023

Review of Large Vision Models and Visual Prompt Engineering

Jiaqi Wang, Zhengliang Liu, Lin Zhao +18

Visual prompt engineering is a fundamental technology in the field of visual and image Artificial General Intelligence, serving as a key component for achieving zero-shot capabilit…

cs.CV2023

SAMAug: Point Prompt Augmentation for Segment Anything Model

Haixing Dai, Chong Ma, Zhiling Yan +14

This paper introduces SAMAug, a novel visual point augmentation method for the Segment Anything Model (SAM) that enhances interactive image segmentation performance. SAMAug generat…

cs.CV20222 cited

BI AVAN: Brain inspired Adversarial Visual Attention Network

Heng Huang, Lin Zhao, Xintao Hu +4

Visual attention is a fundamental mechanism in the human brain, and it inspires the design of attention mechanisms in deep neural networks. However, most of the visual attention st…

cs.CV20221 cited

Eye-gaze-guided Vision Transformer for Rectifying Shortcut Learning

Chong Ma, Lin Zhao, Yuzhong Chen +15

Learning harmful shortcuts such as spurious correlations and biases prevents deep neural networks from learning the meaningful and useful representations, thus jeopardizing the gen…

cs.CV202210 cited

Mask-guided Vision Transformer (MG-ViT) for Few-Shot Learning

Yuzhong Chen, Zhenxiang Xiao, Lin Zhao +10

Learning with little data is challenging but often inevitable in various application scenarios where the labeled data is limited and costly. Recently, few-shot learning (FSL) gaine…