10 citations · 10 across the 3 of their papers we have counts for
2 papers
cs.CV2024
Parameter-Efficient Fine-Tuning Medical Multimodal Large Language Models for Medical Visual Grounding
Jinlong He, Pengfei Li, Gang Liu +1
Multimodal Large Language Models (MLLMs) inherit the superior text understanding capabilities of LLMs and extend these capabilities to multimodal scenarios. These models achieve ex…
cs.CV2022★ 10 cited
Self-supervised vision-language pretraining for Medical visual question answering
Pengfei Li, Gang Liu, Lin Tan +2
Medical image visual question answering (VQA) is a task to answer clinical questions, given a radiographic image, which is a challenging problem that requires a model to integrate…