8 citations · 14 across the 3 of their papers we have counts for
3 papers
cs.CV2025★ 5 cited
Text to Image for Multi-Label Image Recognition with Joint Prompt-Adapter Learning
Chun-Mei Feng, Kai Yu, Xinxing Xu +4
Benefited from image-text contrastive learning, pre-trained vision-language models, e.g., CLIP, allow to direct leverage texts as images (TaI) for parameter-efficient fine-tuning (…
cs.GR2025★ 1 cited
SRDiffusion: Accelerate Video Diffusion Inference via Sketching-Rendering Cooperation
Shenggan Cheng, Yuanxin Wei, Lansong Diao +8
Leveraging the diffusion transformer (DiT) architecture, models like Sora, CogVideoX and Wan have achieved remarkable progress in text-to-video, image-to-video, and video editing t…
cs.CV2024★ 8 cited
Class Balance Matters to Active Class-Incremental Learning
Zitong Huang, Ze Chen, Yuanze Li +6
Few-Shot Class-Incremental Learning has shown remarkable efficacy in efficient learning new concepts with limited annotations. Nevertheless, the heuristic few-shot annotations may…