2 citations · 4 across the 13 of their papers we have counts for
4 papers · 2 filters
VidToMe: Video Token Merging for Zero-Shot Video Editing
Xirui Li, Chao Ma, Xiaokang Yang +1
Diffusion models have made significant advances in generating high-quality images, but their application to video generation has remained challenging due to the complexity of tempo…
Domain Prompt Learning with Quaternion Networks
Qinglong Cao, Zhengqin Xu, Yuntian Chen +2
Prompt learning has emerged as an effective and data-efficient technique in large Vision-Language Models (VLMs). However, when adapting VLMs to specialized domains such as remote s…
Domain-Controlled Prompt Learning
Qinglong Cao, Zhengqin Xu, Yuntian Chen +2
Large pre-trained vision-language models, such as CLIP, have shown remarkable generalization capabilities across various tasks when appropriate text prompts are provided. However,…
Reflection Invariance Learning for Few-shot Semantic Segmentation
Qinglong Cao, Yuntian Chen, Chao Ma +1
Few-shot semantic segmentation (FSS) aims to segment objects of unseen classes in query images with only a few annotated support images. Existing FSS algorithms typically focus on…