1 paper · 1 filter
Hubert Baniecki, Maximilian Muschalik, Fabian Fumagalli +3
Language-image pre-training (LIP) enables the development of vision-language models capable of zero-shot classification, localization, multimodal retrieval, and semantic understand…