2 papers
cs.CV2024
Prompt Generation Networks for Input-Space Adaptation of Frozen Vision Transformers
Jochem Loedeman, Maarten C. Stol, Tengda Han +1
With the introduction of the transformer architecture in computer vision, increasing model scale has been demonstrated as a clear path to achieving performance and robustness gains…
cs.CV2024
Learning to Count without Annotations
Lukas Knobel, Tengda Han, Yuki M. Asano
While recent supervised methods for reference-based object counting continue to improve the performance on benchmark datasets, they have to rely on small datasets due to the cost a…