6 papers
What Does Your Short-Answer VQA Score Actually Measure? Evaluator-Dependent Instability in Multimodal Short-Answer Benchmarks
Guanhua Ye, Niu Jingbin, Yan Li +4
Short-answer VQA benchmarks conflate two distinct quantities: whether a model's answer is semantically correct, and whether that answer matches the surface form expected by the aut…
Beyond Metadata: CAPRA for Hidden Subgroup Analysis under Missing Metadata in Medical Imaging
Yawen Li, Yan Li, Zhe Xue +3
Medical imaging models are often deployed without the demographic, acquisition, and quality metadata needed for subgroup auditing. Once those metadata disappear, clinically critica…
Relation Extraction Model Based on Semantic Enhancement Mechanism
Peiyu Liu, Junping Du, Yingxia Shao +1
Relational extraction is one of the basic tasks related to information extraction in the field of natural language processing, and is an important link and core task in the fields…
Entity Alignment Method of Science and Technology Patent based on Graph Convolution Network and Information Fusion
Runze Fang, Yawen Li, Yingxia Shao +2
The entity alignment of science and technology patents aims to link the equivalent entities in the knowledge graph of different science and technology patent data sources. Most ent…
Trust-free Personalized Decentralized Learning
Yawen Li, Yan Li, Junping Du +3
Personalized collaborative learning in federated settings faces a critical trade-off between customization and participant trust. Existing approaches typically rely on centralized…
ELMM: Efficient Lightweight Multimodal Large Language Models for Multimodal Knowledge Graph Completion
Wei Huang, Peining Li, Meiyu Liang +7
Multimodal Knowledge Graphs (MKGs) extend traditional knowledge graphs by incorporating visual and textual modalities, enabling richer and more expressive entity representations. H…