2 papers
cs.CV2025
Proactive Reasoning-with-Retrieval Framework for Medical Multimodal Large Language Models
Lehan Wang, Yi Qin, Honglong Yang +1
Incentivizing the reasoning ability of Multimodal Large Language Models (MLLMs) is essential for medical applications to transparently analyze medical scans and provide reliable di…
cs.CV2024
FITA: Fine-grained Image-Text Aligner for Radiology Report Generation
Honglong Yang, Hui Tang, Xiaomeng Li
Radiology report generation aims to automatically generate detailed and coherent descriptive reports alongside radiology images. Previous work mainly focused on refining fine-grain…