12 citations · 13 across the 2 of their papers we have counts for
4 papers
WARM-CAT: Warm-Started Test-Time Comprehensive Knowledge Accumulation for Compositional Zero-Shot Learning
Xudong Yan, Songhe Feng, Jiaxin Wang +2
Compositional Zero-Shot Learning (CZSL) aims to recognize novel attribute-object compositions based on the knowledge learned from seen ones. Existing methods suffer from performanc…
MViR: Multi-View Visual-Semantic Representation for Fake News Detection
Haochen Liang, Xinqi Su, Jun Wang +2
With the rise of online social networks, detecting fake news accurately is essential for a healthy online environment. While existing methods have advanced multimodal fake news det…
Towards Deconfounded Image-Text Matching with Causal Inference
Wenhui Li, Xinqi Su, Dan Song +3
Prior image-text matching methods have shown remarkable performance on many benchmark datasets, but most of them overlook the bias in the dataset, which exists in intra-modal and i…
TRRG: Towards Truthful Radiology Report Generation With Cross-modal Disease Clue Enhanced Large Language Model
Yuhao Wang, Chao Hao, Yawen Cui +4
The vision-language modeling capability of multi-modal large language models has attracted wide attention from the community. However, in medical domain, radiology report generatio…