From the 1 of 12 linked papers with an AI index.
12 papers
A report-grounded vision-language foundation model for colonoscopy from 280000 routine reports
Jia Yu, Yan Zhu, Yili He +12
The paper presents EndoCLIP, a vision‑language foundation model for colonoscopy that learns from lesion‑level image‑text pairs extracted from routine colonoscopy reports, achieving…
Dementia Etiology Diagnosis via Collaborative Meta Knowledge Enhancement
Siyuan Du, Mengxi Chen, Xinyang Jiang +6
Although artificial intelligence (AI) has shown promising performance in several medical tasks, accurate dementia etiology diagnosis with AI remains challenging due to complex over…
Dual-Branch Cross-Projection Debiasing through Diffusion-based Disentanglement
Xiangqian Zhao, Xinyang Jiang, Zhipeng Xu +5
Foundation models trained on biased datasets often rely on spurious correlations between target labels and non-causal attributes, resulting in poor generalization on minority group…
Adversarial Domain Prompt Tuning and Generation for Single Domain Generalization
Zhipeng Xu, De Cheng, Xinyang Jiang +3
Single domain generalization (SDG) aims to learn a robust model, which could perform well on many unseen domains while there is only one single domain available for training. One o…
Prompt Disentanglement via Language Guidance and Representation Alignment for Domain Generalization
De Cheng, Zhipeng Xu, Xinyang Jiang +3
Domain Generalization (DG) seeks to develop a versatile model capable of performing effectively on unseen target domains. Notably, recent advances in pre-trained Visual Foundation…
MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents
Jia Yu, Zilong Wang, Xinyang Jiang +2
Computer-use agents could automate repetitive screen-based clinical work, but their reliability in medical graphical user interfaces remains largely unvalidated. Existing benchmark…