1 paper
Tingyu Qu, Tinne Tuytelaars, Marie-Francine Moens
We revisit the weakly supervised cross-modal face-name alignment task; that is, given an image and a caption, we label the faces in the image with the names occurring in the captio…