works on

From the 1 of 7 linked papers with an AI index.

activity
20242026
most citedDepth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR

1 citations · 1 across the 2 of their papers we have counts for

collaborators

7 papers

cs.CV2026

Causal-Adversarial Probing of Clinical Covariates for Prostate MRI Grading

Yipei Wang, Shiqi Huang, Wen Yan +6

The paper introduces an adversarial causal‑reasoning framework to identify which clinical covariates help or hinder deep‑learning models for prostate MRI cancer grading, showing th…

cs.CV20261 cited

Depth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR

Hanyuan Zhang, Lucas He, Runlong He +6

Augmented reality (AR) guidance in laparoscopic liver surgery requires accurate registration of preoperative 3D models to intraoperative 2D video, but remains challenging due to pa…

cs.CV2026

Maximizing T2-Only Prostate Cancer Localization from Expected Diffusion Weighted Imaging

Weixi Yi, Yipei Wang, Wen Yan +10

Multiparametric MRI is increasingly recommended as a first-line noninvasive approach to detect and localize prostate cancer, requiring at minimum diffusion-weighted (DWI) and T2-we…

cs.CV2026

ProFound: A moderate-sized vision foundation model for multi-task prostate imaging

Yipei Wang, Yinsong Xu, Weixi Yi +11

Many diagnostic and therapeutic clinical tasks for prostate cancer increasingly rely on multi-parametric MRI. Automating these tasks is challenging because they necessitate expert…

eess.IV2025

A versatile foundation model for cine cardiac magnetic resonance image analysis tasks

Yunguan Fu, Wenjia Bai, Weixi Yi +7

Here we present a versatile foundation model that can perform a range of clinically-relevant image analysis tasks, including segmentation, landmark localisation, diagnosis, and pro…

cs.CV2025

Analysis of Image-and-Text Uncertainty Propagation in Multimodal Large Language Models with Cardiac MR-Based Applications

Yucheng Tang, Yunguan Fu, Weixi Yi +4

Multimodal large language models (MLLMs) can process and integrate information from multimodality sources, such as text and images. However, interrelationship among input modalitie…