From the 1 of 4 linked papers with an AI index.
4 papers
Benchmarking Foundation and Large Language Models for Few-Shot Medical Image Segmentation
Jinghong Liu, Yuchuan Deng, Fanping Liu +2
The paper presents FAME, a unified benchmark for evaluating few-shot medical image segmentation methods across multiple anatomical sites, imaging modalities, and settings, and anal…
DAPL: Integration of Positive and Negative Descriptions in Text-Based Person Search
Yuchuan Deng, Zhanpeng Hu, Zijie Xin +2
Text-based person search (TBPS) aims to retrieve specific images of individuals from large datasets using textual descriptions. Existing TBPS methods focus primarily on identifying…
Fundus-R1: Training a Fundus-Reading MLLM with Knowledge-Aware Reasoning on Public Data
Yuchuan Deng, Qijie Wei, Kaiheng Qian +6
Fundus imaging such as CFP, OCT and UWF is crucial for the early detection of retinal anomalies and diseases. Fundus image understanding, due to its knowledge-intensive nature, pos…
Empowering Small VLMs to Think with Dynamic Memorization and Exploration
Jiazhen Liu, Yuchuan Deng, Long Chen
Small-scale Vision-Language Models (SVLMs) are exceptionally well-suited for proprietary tasks. Equipping them with thinking capabilities is a critical step to enhance their perfor…