cross-modal learning 1depth perception 1multimodal fusion 1representation learning 1unsupervised semantic segmentation 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CV2026
UMSS: Towards Unsupervised Multi-modal Semantic Segmentation
Haitian Zhang, Thai Duy Nguyen, Xiangyuan Wang +2
The paper introduces UniM2, an unsupervised framework for multimodal semantic segmentation that learns a shared latent space across sensors using cross‑modal correspondence and a h…
cs.CV2025
Frequency-Adaptive Low-Latency Object Detection Using Events and Frames
Haitian Zhang, Xiangyuan Wang, Chang Xu +5
Fusing Events and RGB images for object detection leverages the robustness of Event cameras in adverse environments and the rich semantic information provided by RGB cameras. Howev…