1 paper · 1 filter
Bianca-Mihaela Ganescu, Suchir Salhan, Andrew Caines +1
Training vision-language models on cognitively-plausible amounts of data requires rethinking how models integrate multimodal information. Within the constraints of the Vision track…