2 papers
cs.CV2026
Learning from Noisy Prompts: Saliency-Guided Prompt Distillation for Robust Segmentation with SAM
Jingxuan Kang, Ziqi Zhang, Shaoming Zheng +9
Segmentation is central to clinical diagnosis and monitoring, yet the reliability of modern foundation models in medical imaging still depends on the availability of precise prompt…
cs.CV2025
From Local Details to Global Context: Advancing Vision-Language Models with Attention-Based Selection
Lincan Cai, Jingxuan Kang, Shuang Li +4
Pretrained vision-language models (VLMs), e.g., CLIP, demonstrate impressive zero-shot capabilities on downstream tasks. Prior research highlights the crucial role of visual augmen…