3 citations · 3 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023★ 3 cited
Understanding Segment Anything Model: SAM is Biased Towards Texture Rather than Shape
Chaoning Zhang, Yu Qiao, Shehbaz Tariq +5
In contrast to the human vision that mainly depends on the shape for recognizing the objects, deep image recognition models are widely known to be biased toward texture. Recently,…
cs.CV2023
Toward a Deeper Understanding: RetNet Viewed through Convolution
Chenghao Li, Chaoning Zhang
The success of Vision Transformer (ViT) has been widely reported on a wide range of image recognition tasks. ViT can learn global dependencies superior to CNN, yet CNN's inherent l…