1 paper · 1 filter
Gregory Yauney, Jack Hessel, David Mimno
Images can give us insights into the contextual meanings of words, but current image-text grounding approaches require detailed annotations. Such granular annotation is rare, expen…