4 papers
Unifying Active Learning and Semi-Supervised Learning for Medical Image Segmentation
Bahram Jafrasteh, Cheng Wan, Heejong Kim +2
In practical settings, medical image segmentation models are often developed with limited annotated data rather than fully labeled datasets. Training frequently begins in ultra-low…
Pixel-Space Diffusion Transformers
Renye Yan, Jikang Cheng, You Wu +9
Latent diffusion models (LDMs) enable efficient high-resolution image synthesis by denoising in a VAE-compressed latent space. However, fixed visual tokenizers can discard fine tex…
Anatomically Guided Latent Diffusion for Brain MRI Progression Modeling
Cheng Wan, Bahram Jafrasteh, Ehsan Adeli +2
Accurately modeling longitudinal brain MRI progression is crucial for understanding neurodegenerative diseases and predicting individualized structural changes. Existing state-of-t…
PRISM-Bench: A Benchmark of Puzzle-Based Visual Tasks with CoT Error Detection
Yusu Qian, Cheng Wan, Chao Jia +3
Multimodal large language models (MLLMs) have achieved remarkable progress on vision-language tasks, yet their reasoning processes remain sometimes unreliable. We introduce PRISM-B…