activity
20212024
most citedAutoCLIP: Auto-tuning Zero-Shot Classifiers for Vision-Language Models

1 citations · 1 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CV2024

Attention Is All You Need For Mixture-of-Depths Routing

Advait Gadhikar, Souptik Kumar Majumdar, Niclas Popp +3

Advancements in deep learning are driven by training models with increasingly larger numbers of parameters, which in turn heightens the computational demands. To address this issue…

cs.CV2023

Zero-Shot Visual Classification with Guided Cropping

Piyapat Saranrittichai, Mauricio Munoz, Volker Fischer +1

Pretrained vision-language models, such as CLIP, show promising zero-shot performance across a wide variety of datasets. For closed-set classification tasks, however, there is an i…

cs.CV2023★ 1 cited

AutoCLIP: Auto-tuning Zero-Shot Classifiers for Vision-Language Models

Jan Hendrik Metzen, Piyapat Saranrittichai, Chaithanya Kumar Mummadi

Classifiers built upon vision-language models such as CLIP have shown remarkable zero-shot performance across a broad range of image classification tasks. Prior work has studied di…

cs.CV2022

Multi-Attribute Open Set Recognition

Piyapat Saranrittichai, Chaithanya Kumar Mummadi, Claudia Blaiotta +2

Open Set Recognition (OSR) extends image classification to an open-world setting, by simultaneously classifying known classes and identifying unknown ones. While conventional OSR a…

cs.CV2022

Overcoming Shortcut Learning in a Target Domain by Generalizing Basic Visual Factors from a Source Domain

Piyapat Saranrittichai, Chaithanya Kumar Mummadi, Claudia Blaiotta +2

Shortcut learning occurs when a deep neural network overly relies on spurious correlations in the training dataset in order to solve downstream tasks. Prior works have shown how th…

cs.CV2021

DiagViB-6: A Diagnostic Benchmark Suite for Vision Models in the Presence of Shortcut and Generalization Opportunities

Elias Eulig, Piyapat Saranrittichai, Chaithanya Kumar Mummadi +4

Common deep neural networks (DNNs) for image classification have been shown to rely on shortcut opportunities (SO) in the form of predictive and easy-to-represent visual factors. T…