2 papers
cs.CV2026
UniCorrn: Unified Correspondence Transformer Across 2D and 3D
Prajnan Goswami, Tianye Ding, Feng Liu +1
Visual correspondence across image-to-image (2D-2D), image-to-point cloud (2D-3D), and point cloud-to-point cloud (3D-3D) geometric matching forms the foundation for numerous 3D vi…
cs.CV2025
DynRefer: Delving into Region-level Multimodal Tasks via Dynamic Resolution
Yuzhong Zhao, Feng Liu, Yue Liu +4
One fundamental task of multimodal models is to translate referred image regions to human preferred language descriptions. Existing methods, however, ignore the resolution adaptabi…