Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Knowledge to Sight: Reasoning over Visual Attributes via Knowledge Decomposition for Abnormality Grounding
Jun Li, Che Liu, Wenjia Bai +4
In this work, we address the problem of grounding abnormalities in medical images, where the goal is to localize clinical findings based on textual descriptions. While generalist V…
cs.CV2025
Enhancing Abnormality Grounding for Vision Language Models with Knowledge Descriptions
Jun Li, Che Liu, Wenjia Bai +3
Visual Language Models (VLMs) have demonstrated impressive capabilities in visual grounding tasks. However, their effectiveness in the medical domain, particularly for abnormality…
cs.CV2024
FMBench: Benchmarking Fairness in Multimodal Large Language Models on Medical Tasks
Peiran Wu, Che Liu, Canyu Chen +3
Advancements in Multimodal Large Language Models (MLLMs) have significantly improved medical task performance, such as Visual Question Answering (VQA) and Report Generation (RG). H…