5 citations · 5 across the 2 of their papers we have counts for
5 papers
Image Captioning with Context-Aware Auxiliary Guidance
Zeliang Song, Xiaofei Zhou, Zhendong Mao +1
Image captioning is a challenging computer vision task, which aims to generate a natural language description of an image. Most recent researches follow the encoder-decoder framewo…
Cross-modal Knowledge Reasoning for Knowledge-based Visual Question Answering
Jing Yu, Zihao Zhu, Yujing Wang +3
Knowledge-based Visual Question Answering (KVQA) requires external knowledge beyond the visible content to answer questions about an image. This ability is challenging but indispen…
RLINK: Deep Reinforcement Learning for User Identity Linkage
Xiaoxue Li, Yanan Cao, Yanmin Shang +3
User identity linkage is a task of recognizing the identities of the same user across different social networks (SN). Previous works tackle this problem via estimating the pairwise…
Deep Active Learning for Anchor User Prediction
Anfeng Cheng, Chuan Zhou, Hong Yang +4
Predicting pairs of anchor users plays an important role in the cross-network analysis. Due to the expensive costs of labeling anchor users for training prediction models, we consi…
Modeling Text with Graph Convolutional Network for Cross-Modal Information Retrieval
Jing Yu, Yuhang Lu, Zengchang Qin +4
Cross-modal information retrieval aims to find heterogeneous data of various modalities from a given query of one modality. The main challenge is to map different modalities into a…