3 papers
cs.RO2026
OMNI-PoseX: A Fast Vision Model for 6D Object Pose Estimation in Embodied Tasks
Michael Zhang, Wei Ying, Fangwen Chen +2
Accurate 6D object pose estimation is a fundamental capability for embodied agents, yet remains highly challenging in open-world environments. Many existing methods often rely on c…
cs.CV2025
CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images
Bin Hu, Chenqiang Gao, Shurui Liu +4
Image translation is one of the crucial approaches for mitigating information deficiencies in the infrared and visible modalities, while also facilitating the enhancement of modali…
cs.CV2024
IVGF: The Fusion-Guided Infrared and Visible General Framework
Fangcen Liu, Chenqiang Gao, Fang Chen +3
Infrared and visible dual-modality tasks such as semantic segmentation and object detection can achieve robust performance even in extreme scenes by fusing complementary informatio…