32 citations · 35 across the 6 of their papers we have counts for
3 papers · 1 filter
Common Data Properties Limit Object-Attribute Binding in CLIP
Bijay Gurung, David T. Hoffmann, Thomas Brox
Contrastive vision-language models like CLIP are used for a large variety of applications, such as zero-shot classification or as vision encoder for multi-modal models. Despite the…
Floxels: Fast Unsupervised Voxel Based Scene Flow Estimation
David T. Hoffmann, Syed Haseeb Raza, Hanqiu Jiang +3
Scene flow estimation is a foundational task for many robotic applications, including robust dynamic object detection, automatic labeling, and sensor synchronization. Two types of…
Unlocking In-Context Learning for Natural Datasets Beyond Language Modelling
Jelena Bratulić, Sudhanshu Mittal, David T. Hoffmann +5
Large Language Models (LLMs) exhibit In-Context Learning (ICL), which enables the model to perform new tasks conditioning only on the examples provided in the context without updat…