collaborators

41 papers

cs.CV2026

Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching

Daniele Molino, Alessio Zoboli, Camillo Maria Caruso +2

Cross-modality medical image translation can reduce the burden of multi-modal acquisitions, yet the field remains constrained by two coupled limitations: methods operate on 2D slic…

cs.LG2026

Is Self-Pretraining really useful to improve diagnosis in medical Time Series?

Omar Coser, Antonio Orvieto, Paolo Soda +1

Inspired by recent evidence that transformer architectures benefit from Self-PreTraining (SPT) on long-context benchmarks, we investigate whether similar gains extend to multimodal…

cs.CV2026

SHOVIR: A Benchmark for Evaluating Vision Shortcut Learning in Radiology Report Generation

Filippo Ruffini, Marco Salmé, Rosa Sicilia +2

Current evaluation protocols for Vision-Language Models (VLMs) in Radiology Report Generation (RRG) rely on report-level metrics that measure lexical overlap or aggregate clinical…

cs.LG2026

Probabilistic NDVI Forecasting from Sparse Satellite Time Series and Weather Covariates

Irene Iele, Giulia Romoli, Daniele Molino +4

Short-term forecasting of vegetation dynamics is a key enabler for data-driven decision support in precision agriculture. Normalized Difference Vegetation Index (NDVI) forecasting…

cs.LG2026

VegSim: A Geospatial World Model for Scenario-Conditioned Vegetation Simulation

Irene Iele, Elena Mulero Ayllón, Paolo Soda +1

Vegetation monitoring under climate stress requires answering not only how it will evolve given the expected weather, but how it would respond to alternative meteorological conditi…

cs.CV2026

Retrieval-Augmented Anatomical Guidance for Text-to-CT Generation

Daniele Molino, Camillo Maria Caruso, Paolo Soda +1

Text-conditioned generative models for volumetric medical imaging provide semantic control but lack explicit anatomical guidance, often resulting in outputs that are spatially ambi…