1 paper
Yangtao Wang, Xi Shen, Shell Hu +3
Transformers trained with self-supervised learning using self-distillation loss (DINO) have been shown to produce attention maps that highlight salient foreground objects. In this…