2 papers
cs.LG2026
UVLM: A Universal Vision-Language Model Loader for Reproducible Multimodal Benchmarking
Joan Perez, Giovanni Fusco
Vision-Language Models (VLMs) have emerged as powerful tools for image understanding tasks, yet their practical deployment remains hindered by significant architectural heterogenei…
cs.CV2025
Streetscape Analysis with Generative AI (SAGAI): Vision-Language Assessment and Mapping of Urban Scenes
Joan Perez, Giovanni Fusco
Streetscapes are an essential component of urban space. Their assessment is presently either limited to morphometric properties of their mass skeleton or requires labor-intensive q…