3 papers
cs.CV2019
Learning Visual Relation Priors for Image-Text Matching and Image Captioning with Neural Scene Graph Generators
Kuang-Huei Lee, Hamid Palangi, Xi Chen +2
Grounding language to visual relations is critical to various language-and-vision applications. In this work, we tackle two fundamental language-and-vision tasks: image-text matchi…
cs.CV2018
Stacked Cross Attention for Image-Text Matching
Kuang-Huei Lee, Xi Chen, Gang Hua +2
In this paper, we study the problem of image-text matching. Inferring the latent semantic alignment between objects or other salient stuff (e.g. snow, sky, lawn) and the correspond…
cs.CV2018
Web-Scale Responsive Visual Search at Bing
Houdong Hu, Yan Wang, Linjun Yang +7
In this paper, we introduce a web-scale general visual search system deployed in Microsoft Bing. The system accommodates tens of billions of images in the index, with thousands of…