works on

From the 2 of 5 linked papers with an AI index.

collaborators

5 papers

eess.IV2026

TCAM-Diff: Triplane-Aware Cross-Attention Medical Diffusion Model

Zhenkai Zhang, Krista A. Ehinger, Tom Drummond

The paper presents TCAM-Diff, a diffusion-based model that uses a triplane-aware cross‑attention mechanism and a decoder‑only autoencoder to efficiently generate high‑resolution 3D…

eess.IV2026

From Sparse X-rays to 3D CT: Training-Free Reconstruction with Diffusion Priors

Zhenkai Zhang, Markus Hiller, Krista A. Ehinger +1

The paper introduces TF-PRDiT, a training‑free framework that uses a frozen 3D diffusion transformer prior to reconstruct CT volumes from sparse X‑ray projections, and can be appli…

cs.CV2026

Pixel-Level Residual Diffusion Transformer: Scalable 3D CT Volume Generation

Zhenkai Zhang, Markus Hiller, Krista A. Ehinger +1

Generating high-resolution 3D CT volumes with fine details remains challenging due to substantial computational demands and optimization difficulties inherent to existing generativ…

cs.LG2026

Improving Denoising Diffusion Models via Simultaneous Estimation of Image and Noise

Zhenkai Zhang, Krista A. Ehinger, Tom Drummond

This paper introduces two key contributions aimed at improving the speed and quality of images generated through inverse diffusion processes. The first contribution involves repara…

cs.CV2025

Reasoning Like Experts: Leveraging Multimodal Large Language Models for Drawing-based Psychoanalysis

Xueqi Ma, Yanbei Jiang, Sarah Erfani +4

Multimodal Large Language Models (MLLMs) have demonstrated exceptional performance across various objective multimodal perception tasks, yet their application to subjective, emotio…