diffusion models 23d image generation 1cross-attention 1CT reconstruction 1inverse problems 1medical imaging 1training-free methods 1triplane representation 1volumetric imaging 1
From the 2 of 6 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Pixel-Level Residual Diffusion Transformer: Scalable 3D CT Volume Generation
Zhenkai Zhang, Markus Hiller, Krista A. Ehinger +1
Generating high-resolution 3D CT volumes with fine details remains challenging due to substantial computational demands and optimization difficulties inherent to existing generativ…
cs.CV2025
Reasoning Like Experts: Leveraging Multimodal Large Language Models for Drawing-based Psychoanalysis
Xueqi Ma, Yanbei Jiang, Sarah Erfani +4
Multimodal Large Language Models (MLLMs) have demonstrated exceptional performance across various objective multimodal perception tasks, yet their application to subjective, emotio…
cs.CV2024
Perceiving Longer Sequences With Bi-Directional Cross-Attention Transformers
Markus Hiller, Krista A. Ehinger, Tom Drummond
We present a novel bi-directional Transformer architecture (BiXT) which scales linearly with input size in terms of computational cost and memory consumption, but does not suffer t…