1 paper
Bryan A. Plummer, Paige Kordas, M. Hadi Kiapour +3
This paper presents an approach for grounding phrases in images which jointly learns multiple text-conditioned embeddings in a single end-to-end model. In order to differentiate te…