Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
UniCorn: Towards Self-Improving Unified Multimodal Models through Self-Generated Supervision
Ruiyan Han, Zhen Fang, XinYu Sun +9
While Unified Multimodal Models (UMMs) have achieved remarkable success in cross-modal comprehension, a significant gap persists in their ability to leverage such internal knowledg…
cs.CV2023
RGM: A Robust Generalizable Matching Model
Songyan Zhang, Xinyu Sun, Hao Chen +2
Finding corresponding pixels within a pair of images is a fundamental computer vision task with various applications. Due to the specific requirements of different tasks like optic…