81 citations · 112 across the 8 of their papers we have counts for
1 paper · 1 filter
Jheng-Hong Yang, Carlos Lassance, Rafael Sampaio de Rezende +4
This paper presents the AToMiC (Authoring Tools for Multimedia Content) dataset, designed to advance research in image/text cross-modal retrieval. While vision-language pretrained…