41 papers
Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching
Daniele Molino, Alessio Zoboli, Camillo Maria Caruso +2
Cross-modality medical image translation can reduce the burden of multi-modal acquisitions, yet the field remains constrained by two coupled limitations: methods operate on 2D slic…
Is Self-Pretraining really useful to improve diagnosis in medical Time Series?
Omar Coser, Antonio Orvieto, Paolo Soda +1
Inspired by recent evidence that transformer architectures benefit from Self-PreTraining (SPT) on long-context benchmarks, we investigate whether similar gains extend to multimodal…
SHOVIR: A Benchmark for Evaluating Vision Shortcut Learning in Radiology Report Generation
Filippo Ruffini, Marco Salmé, Rosa Sicilia +2
Current evaluation protocols for Vision-Language Models (VLMs) in Radiology Report Generation (RRG) rely on report-level metrics that measure lexical overlap or aggregate clinical…
Probabilistic NDVI Forecasting from Sparse Satellite Time Series and Weather Covariates
Irene Iele, Giulia Romoli, Daniele Molino +4
Short-term forecasting of vegetation dynamics is a key enabler for data-driven decision support in precision agriculture. Normalized Difference Vegetation Index (NDVI) forecasting…
VegSim: A Geospatial World Model for Scenario-Conditioned Vegetation Simulation
Irene Iele, Elena Mulero Ayllón, Paolo Soda +1
Vegetation monitoring under climate stress requires answering not only how it will evolve given the expected weather, but how it would respond to alternative meteorological conditi…
Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation
Daniele Molino, Camillo Maria Caruso, Paolo Soda +1
Text-conditioned generative models for volumetric medical imaging provide semantic control but lack explicit anatomical guidance, often resulting in outputs that are spatially ambi…