2 papers
cs.CV2026
Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA
Yuetian Du, Yucheng Wang, Ming Kong +4
Multimodal Large Language Models (MLLMs) show great potential in medical tasks, but their elicited confidence often misaligns with actual accuracy, potentially leading to misdiagno…
cs.CV2025
Style-Aligned Image Composition for Robust Detection of Abnormal Cells in Cytopathology
Qiuyi Qi, Xin Li, Ming Kong +4
Challenges such as the lack of high-quality annotations, long-tailed data distributions, and inconsistent staining styles pose significant obstacles to training neural networks to…