50 citations · 80 across the 5 of their papers we have counts for
4 papers · 1 filter
LPF: A Language-Prior Feedback Objective Function for De-biased Visual Question Answering
Zujie Liang, Haifeng Hu, Jiaying Zhu
Most existing Visual Question Answering (VQA) systems tend to overly rely on language bias and hence fail to reason from the visual clue. To address this issue, we propose a novel…
Universal Multi-Source Domain Adaptation
Yueming Yin, Zhen Yang, Haifeng Hu +1
Unsupervised domain adaptation enables intelligent models to transfer knowledge from a labeled source domain to a similar but unlabeled target domain. Recent study reveals that kno…
Adaptive Interaction Modeling via Graph Operations Search
Haoxin Li, Wei-Shi Zheng, Yu Tao +2
Interaction modeling is important for video action analysis. Recently, several works design specific structures to model interactions in videos. However, their structures are manua…
Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal Fusion
Sijie Mai, Haifeng Hu, Songlong Xing
Learning joint embedding space for various modalities is of vital importance for multimodal fusion. Mainstream modality fusion approaches fail to achieve this goal, leaving a modal…