2 papers
cs.CV2024
WATT: Weight Average Test-Time Adaptation of CLIP
David Osowiechi, Mehrdad Noori, Gustavo Adolfo Vargas Hakim +7
Vision-Language Models (VLMs) such as CLIP have yielded unprecedented performance for zero-shot image classification, yet their generalization capability may still be seriously cha…
cs.CV2024
NC-TTT: A Noise Contrastive Approach for Test-Time Training
David Osowiechi, Gustavo A. Vargas Hakim, Mehrdad Noori +5
Despite their exceptional performance in vision tasks, deep learning models often struggle when faced with domain shifts during testing. Test-Time Training (TTT) methods have recen…