Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Learning Visual Generative Priors without Text
Shuailei Ma, Kecheng Zheng, Ying Wei +7
Although text-to-image (T2I) models have recently thrived as visual generative priors, their reliance on high-quality text-image pairs makes scaling up expensive. We argue that gra…
cs.CV2024
SKDF: A Simple Knowledge Distillation Framework for Distilling Open-Vocabulary Knowledge to Open-world Object Detector
Shuailei Ma, Yuefeng Wang, Ying Wei +4
In this paper, we attempt to specialize the VLM model for OWOD tasks by distilling its open-world knowledge into a language-agnostic detector. Surprisingly, we observe that the com…
cs.CV2024
Understanding the Multi-modal Prompts of the Pre-trained Vision-Language Model
Shuailei Ma, Chen-Wei Xie, Ying Wei +5
Prompt learning has emerged as an efficient alternative for fine-tuning foundational models, such as CLIP, for various downstream tasks. However, there is no work that provides a c…