5 papers
Grounding Free-Form Instructions for Fashion Complementary Image Generation
Matteo Attimonelli, Claudio Pomo, Alessandro De Bellis +3
Fashion complementary image generation (CIG) aims to create garments that stylistically match a seed item based on user intent, making it a natural multimodal grounding problem whe…
Who Are You Explaining To? A Multi-Agent System for Audience-Aware XAI Narratives
Francesco Musicco, Danilo Danese, Giuseppe Fasano +3
Feature-attribution methods such as SHAP provide useful evidence about individual model predictions, but their numerical outputs are rarely sufficient for audiences with different…
Now You Have My Healthy Attention: A U-DiT for Brain-MRI Inpainting
Danilo Danese, Angela Lombardi, Tommaso Di Noia
The ASNR-MICCAI BraTS Local Synthesis (Inpainting) task asks for the anatomically plausible completion of healthy brain tissue within a masked region of a T1-weighted MRI, providin…
Do Recommender Systems Really Leverage Multimodal Content? A Comprehensive Analysis on Multimodal Representations for Recommendation
Claudio Pomo, Matteo Attimonelli, Danilo Danese +2
Multimodal Recommender Systems aim to improve recommendation accuracy by integrating heterogeneous content, such as images and textual metadata. While effective, it remains unclear…
Large-scale Benchmarks for Multimodal Recommendation with Ducho
Matteo Attimonelli, Danilo Danese, Angela Di Fazio +3
The common multimodal recommendation pipeline involves (i) extracting multimodal features, (ii) refining their high-level representations to suit the recommendation task, (iii) opt…