1 paper
Thinesh Thiyakesan Ponbagavathi, Constantin Seibold, Alina Roitberg
Adapting image-pretrained backbones to video typically relies on time-domain adapters tuned to a single temporal scale. Our experiments show that these modules pick up static image…