3 papers
cs.CV2026
Multi-View Mixture-of-Experts with Vision-Language Reranking for Cross-View Object Geo-Localization
Xuyu Fan, Qi Ming, Zhu Han +6
Cross-view object geo-localization (CVOGL) locates a target in satellite imagery using drone or street-view queries. Existing methods train separate detectors for each viewpoint, l…
cs.CV2026
Stealthy in Semantics, Antagonistic in Space: Attacking Visible-Infrared Object Detectors via Object-Level Misalignment
Yueqi Zhu, Qi Ming, Guo Cheng +6
Visible-infrared object detectors are used for robust perception under challenging illumination and weather conditions. Current physical attacks apply conspicuous patches to spatia…
cs.CV2026
Aligned Consensus Teaching for Label-Efficient Oriented Object Detection in Weakly-Aligned Visible-Infrared Imagery
Qi Ming, Xiaxin Yuan, Jiahuan Zhou +6
Visible-infrared object detection (VIOD) detects objects with oriented bounding boxes from paired visible and infrared images. Existing methods depend on costly dual-modality annot…