3 papers
cs.CV2026
RadSight: Towards Perceptually Reliable Multimodal Radiology Image Understanding
Jianqin Liu, Weiwei Cao, Wanxing Chang +7
Medical multimodal large language models (MLLMs) are increasingly expected to perform complex image understanding tasks, yet their reliability is often compromised by frequent erro…
cs.CE2026
AtomiMed: Hierarchical Atomic Fact-Checking for Universal Clinical-Aware Medical Report Evaluation
Yuan Wang, Wanxing Chang, Songtao Jiang +8
Traditional metrics for Medical Report Generation (MRG) predominantly rely on surface-level n-gram overlap, which fails to capture clinical factual accuracy and often overlooks cat…
cs.AI2026
CT-FineBench: A Diagnostic Fidelity Benchmark for Fine-Grained Evaluation of CT Report Generation
Ruifeng Yuan, Wanxing Chang, Weiwei Cao +4
The evaluation of generated reports remains a critical challenge in Computed Tomography (CT) report generation, due to the large volume of text, the diversity and complexity of fin…