2 papers
cs.CV2026
Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models
Siyuan Yang, Jun Liu, Hao Cheng +5
Recent advances in large-scale pretrained vision models have demonstrated impressive capabilities across a wide range of downstream tasks, including cross-modal and multi-modal sce…
cs.CV2026
Modality-Aware Feature Matching in Visual and Vision-Language Applications: A Comprehensive Survey
Weide Liu, Wei Zhou, Jun Liu +4
Feature matching is a cornerstone task in computer vision, essential for applications such as image retrieval, stereo matching, 3D reconstruction, and SLAM. This survey comprehensi…