2 papers
cs.CV2025
EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery
Guankun Wang, Long Bai, Junyi Wang +13
Recently, Multimodal Large Language Models (MLLMs) have demonstrated their immense potential in computer-aided diagnosis and decision-making. In the context of robotic-assisted sur…
cs.CV2025
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
Yifei Zhang, James Song, Siyi Gu +4
Explainable AI (XAI) has gained significant attention for providing insights into the decision-making processes of deep learning models, particularly for image classification tasks…