3 papers
cs.CV2026
VidPrism: Heterogeneous Mixture of Experts for Image-to-Video Transfer
Rui Lin, Chuanming Wang, Huadong Ma
With the rapid development of pre-training technologies, adapting large-scale Vision-Language Models (VLMs) for video understanding \emph{\ie} image-to-video transfer learning has…
cs.CV2025
XR-VLM: Cross-Relationship Modeling with Multi-part Prompts and Visual Features for Fine-Grained Recognition
Chuanming Wang, Henming Mao, Huanhuan Zhang +2
Vision-Language Models (VLMs) have demonstrated impressive performance on various visual tasks, yet they still require adaptation on downstream tasks to achieve optimal performance…
cs.CV2024
Towards Efficient Object Re-Identification with A Novel Cloud-Edge Collaborative Framework
Chuanming Wang, Yuxin Yang, Mengshi Qi +1
Object re-identification (ReID) is committed to searching for objects of the same identity across cameras, and its real-world deployment is gradually increasing. Current ReID metho…