14 citations · 24 across the 10 of their papers we have counts for
1 paper · 2 filters
Hubert Baniecki, Maximilian Muschalik, Fabian Fumagalli +3
Language-image pre-training (LIP) enables the development of vision-language models capable of zero-shot classification, localization, multimodal retrieval, and semantic understand…