17 citations · 17 across the 2 of their papers we have counts for
9 papers
Image Scene Graph Generation (SGG) Benchmark
Xiaotian Han, Jianwei Yang, Houdong Hu +3
There is a surge of interest in image scene graph generation (object, attribute and relationship detection) due to the need of building fine-grained image understanding models that…
Oscar: Object-Semantics Aligned Pre-training for Vision-Language Tasks
Xiujun Li, Xi Yin, Chunyuan Li +9
Large-scale pre-training methods of learning cross-modal representations on image-text pairs are becoming popular for vision-language tasks. While existing methods simply concatena…
Applications of Generative Adversarial Models in Visual Search Reformulation
Kyle Xiao, Houdong Hu, Yan Wang
Query reformulation is the process by which a input search query is refined by the user to match documents outside the original top-n results. On average, roughly 50% of text searc…
Learning Visual Relation Priors for Image-Text Matching and Image Captioning with Neural Scene Graph Generators
Kuang-Huei Lee, Hamid Palangi, Xi Chen +2
Grounding language to visual relations is critical to various language-and-vision applications. In this work, we tackle two fundamental language-and-vision tasks: image-text matchi…
Unified Vision-Language Pre-Training for Image Captioning and VQA
Luowei Zhou, Hamid Palangi, Lei Zhang +3
This paper presents a unified Vision-Language Pre-training (VLP) model. The model is unified in that (1) it can be fine-tuned for either vision-language generation (e.g., image cap…
Symmetry Breaking in Density Functional Theory due to Dirac Exchange for a Hydrogen Molecule
Michael Holst, Houdong Hu, Jianfeng Lu +3
We study symmetry breaking in the mean field solutions to the 2 electron hydrogen molecule within Kohn Sham (KS) local spin density function theory with Dirac exchange (the XLDA mo…