1 paper
Jim Berend, Reduan Achtibat, Daniel Schäffer +4
Vision Transformers (ViTs) are central to most modern vision models, yet obtaining input attributions that are fine-grained, faithful, and stable remains challenging. Layer-wise Re…