1 paper · 1 filter
Moran Yanuka, Morris Alper, Hadar Averbuch-Elor +1
Web-scale training on paired text-image data is becoming increasingly central to multimodal learning, but is challenged by the highly noisy nature of datasets in the wild. Standard…