papers

Publications (95)

cs.CL2026

A Geolocation-Aware Multimodal Approach for Ecological Prediction

Valerie Zermatten, Chiara Vanalli, Gencer Sumbul +2

While integrating multiple modalities has the potential to improve environmental monitoring, current approaches struggle to combine data sources with heterogeneous formats or conte…

cs.CV2019

Semantically Interpretable Activation Maps: what-where-how explanations within CNNs

Diego Marcos, Sylvain Lobry, Devis Tuia

A main issue preventing the use of Convolutional Neural Networks (CNN) in end user applications is the low level of transparency in the decision process. Previous work on CNN inter…

cs.CV2022

Geo-Information Harvesting from Social Media Data

Xiao Xiang Zhu, Yuanyuan Wang, Mrinalini Kochupillai +9

As unconventional sources of geo-information, massive imagery and text messages from open platforms and social media form a temporally quasi-seamless, spatially multi-perspective s…

cs.CY2024

Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants

Beatriz Borges, Negar Foroutan, Deniz Bayazit +87

AI assistants are being increasingly used by students enrolled in higher education institutions. While these tools provide opportunities for improved teaching and education, they a…

cs.LG2025

What to align in multimodal contrastive learning?

Benoit Dufumier, Javiera Castillo-Navarro, Devis Tuia +1

Humans perceive the world through multisensory integration, blending the information of different modalities to adapt their behavior. Contrastive learning offers an appealing solut…

cs.CV2017

Towards seamless multi-view scene analysis from satellite to street-level

Sébastien Lefèvre, Devis Tuia, Jan Dirk Wegner +2

In this paper, we discuss and review how combined multi-view imagery from satellite to street-level can benefit scene analysis. Numerous works exist that merge information from rem…