3 papers
cs.CV2025
Common Data Properties Limit Object-Attribute Binding in CLIP
Bijay Gurung, David T. Hoffmann, Thomas Brox
Contrastive vision-language models like CLIP are used for a large variety of applications, such as zero-shot classification or as vision encoder for multi-modal models. Despite the…
cs.CV2025
Using Knowledge Graphs to harvest datasets for efficient CLIP model training
Simon Ging, Sebastian Walter, Jelena Bratulić +3
Training high-quality CLIP models typically requires enormous datasets, which limits the development of domain-specific models -- especially in areas that even the largest CLIP mod…
cs.CL2025
Unlocking In-Context Learning for Natural Datasets Beyond Language Modelling
Jelena Bratulić, Sudhanshu Mittal, David T. Hoffmann +5
Large Language Models (LLMs) exhibit In-Context Learning (ICL), which enables the model to perform new tasks conditioning only on the examples provided in the context without updat…