3 papers
cs.CV2025
MedSG-Bench: A Benchmark for Medical Image Sequences Grounding
Jingkun Yue, Siqi Zhang, Zinan Jia +4
Visual grounding is essential for precise perception and reasoning in multimodal large language models (MLLMs), especially in medical imaging domains. While existing medical visual…
cs.CV2024
MoVL:Exploring Fusion Strategies for the Domain-Adaptive Application of Pretrained Models in Medical Imaging Tasks
Haijiang Tian, Jingkun Yue, Xiaohong Liu +3
Medical images are often more difficult to acquire than natural images due to the specialism of the equipment and technology, which leads to less medical image datasets. So it is h…
cs.LG2024
SPD-CFL: Stepwise Parameter Dropout for Efficient Continual Federated Learning
Yuning Yang, Han Yu, Chuan Sun +5
Federated Learning (FL) is a collaborative machine learning paradigm for training models on local sensitive data with privacy protection. Pre-trained transformer-based models have…