2 papers
cs.RO2026
OMNI-PoseX: A Fast Vision Model for 6D Object Pose Estimation in Embodied Tasks
Michael Zhang, Wei Ying, Fangwen Chen +2
Accurate 6D object pose estimation is a fundamental capability for embodied agents, yet remains highly challenging in open-world environments. Many existing methods often rely on c…
cs.CV2025
CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images
Bin Hu, Chenqiang Gao, Shurui Liu +4
Image translation is one of the crucial approaches for mitigating information deficiencies in the infrared and visible modalities, while also facilitating the enhancement of modali…