7 citations
1 paper
Shakeeb Murtaza, Soufiane Belharbi, Marco Pedersoli +2
Self-supervised vision transformers (SSTs) have shown great potential to yield rich localization maps that highlight different objects in an image. However, these maps remain class…