Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
GroundSet: A Cadastral-Grounded Dataset for Spatial Understanding with Vector Data
Roger Ferrod, Maël Lecene, Krishna Sapkota +4
Precise spatial understanding in Earth Observation is essential for translating raw aerial imagery into actionable insights for critical applications like urban planning, environme…
cs.CV2025
Revisiting Cross-Modal Knowledge Distillation: A Disentanglement Approach for RGBD Semantic Segmentation
Roger Ferrod, Cássio F. Dantas, Luigi Di Caro +1
Multi-modal RGB and Depth (RGBD) data are predominant in many domains such as robotics, autonomous driving and remote sensing. The combination of these multi-modal data enhances en…
cs.CV2024
Towards a multimodal framework for remote sensing image change retrieval and captioning
Roger Ferrod, Luigi Di Caro, Dino Ienco
Recently, there has been increasing interest in multimodal applications that integrate text with other modalities, such as images, audio and video, to facilitate natural language i…