From the 2 of 5 linked papers with an AI index.
5 papers
TCAM-Diff: Triplane-Aware Cross-Attention Medical Diffusion Model
Zhenkai Zhang, Krista A. Ehinger, Tom Drummond
The paper presents TCAM-Diff, a diffusion-based model that uses a triplane-aware cross‑attention mechanism and a decoder‑only autoencoder to efficiently generate high‑resolution 3D…
From Sparse X-rays to 3D CT: Training-Free Reconstruction with Diffusion Priors
Zhenkai Zhang, Markus Hiller, Krista A. Ehinger +1
The paper introduces TF-PRDiT, a training‑free framework that uses a frozen 3D diffusion transformer prior to reconstruct CT volumes from sparse X‑ray projections, and can be appli…
Pixel-Level Residual Diffusion Transformer: Scalable 3D CT Volume Generation
Zhenkai Zhang, Markus Hiller, Krista A. Ehinger +1
Generating high-resolution 3D CT volumes with fine details remains challenging due to substantial computational demands and optimization difficulties inherent to existing generativ…
Improving Denoising Diffusion Models via Simultaneous Estimation of Image and Noise
Zhenkai Zhang, Krista A. Ehinger, Tom Drummond
This paper introduces two key contributions aimed at improving the speed and quality of images generated through inverse diffusion processes. The first contribution involves repara…
Reasoning Like Experts: Leveraging Multimodal Large Language Models for Drawing-based Psychoanalysis
Xueqi Ma, Yanbei Jiang, Sarah Erfani +4
Multimodal Large Language Models (MLLMs) have demonstrated exceptional performance across various objective multimodal perception tasks, yet their application to subjective, emotio…