3 papers
cs.CV2026
XiDepth: a Lightweight and Efficient Network for Self-supervised Monocular Depth Estimation
Elena Izzo, Riccardo Toniolo, Lamberto Ballan
Self-supervised monocular depth estimation has emerged as an appealing solution to design lightweight and effective models for deployment on computationally constrained devices due…
cs.CV2025
7Bench: a Comprehensive Benchmark for Layout-guided Text-to-image Models
Elena Izzo, Luca Parolari, Davide Vezzaro +1
Layout-guided text-to-image models offer greater control over the generation process by explicitly conditioning image synthesis on the spatial arrangement of elements. As a result,…
cs.CV2024
Harlequin: Color-driven Generation of Synthetic Data for Referring Expression Comprehension
Luca Parolari, Elena Izzo, Lamberto Ballan
Referring Expression Comprehension (REC) aims to identify a particular object in a scene by a natural language expression, and is an important topic in visual language understandin…