activity
20172020
most citedMulti-modal Factorized Bilinear Pooling with Co-Attention Learning for Visual Question Answering

102 citations · 242 across the 14 of their papers we have counts for

collaborators
Showing cs.CVShow all

16 papers · 1 filter

cs.CV20203 cited

Deep Fusion Siamese Network for Automatic Kinship Verification

Jun Yu, Mengyan Li, Xinlong Hao +1

Automatic kinship verification aims to determine whether some individuals belong to the same family. It is of great research significance to help missing persons reunite with their…

cs.CV20201 cited

Retrieval of Family Members Using Siamese Neural Network

Jun Yu, Guochen Xie, Mengyan Li +1

Retrieval of family members in the wild aims at finding family members of the given subject in the dataset, which is useful in finding the lost children and analyzing the kinship.…

cs.CV2020

Deep Multimodal Neural Architecture Search

Zhou Yu, Yuhao Cui, Jun Yu +3

Designing effective neural networks is fundamentally important in deep multimodal learning. Most existing works focus on a single task and design neural architectures manually, whi…

cs.CV201911 cited

ActivityNet-QA: A Dataset for Understanding Complex Web Videos via Question Answering

Zhou Yu, Dejing Xu, Jun Yu +4

Recent developments in modeling language and vision have been successfully applied to image question answering. It is both crucial and natural to extend this research direction to…

cs.CV201930 cited

Multimodal Transformer with Multi-View Visual Representation for Image Captioning

Jun Yu, Jing Li, Zhou Yu +1

Image captioning aims to automatically generate a natural language description of a given image, and most state-of-the-art models have adopted an encoder-decoder framework. The fra…

cs.CV201910 cited

Single Pixel Reconstruction for One-stage Instance Segmentation

Jun Yu, Jinghan Yao, Jian Zhang +2

Object instance segmentation is one of the most fundamental but challenging tasks in computer vision, and it requires the pixel-level image understanding. Most existing approaches…