1 paper
Tanjim Islam Riju, Shuchismita Anwar, Saman Sarker Joy +2
Medical vision-language models still struggle to match radiologists' attention and to verbalize findings with explicit spatial grounding. We address this gap with a two-stage multi…