3 papers
cs.CV2023
LAMM: Label Alignment for Multi-Modal Prompt Learning
Jingsheng Gao, Jiacheng Ruan, Suncheng Xiang +5
With the success of pre-trained visual-language (VL) models such as CLIP in visual representation tasks, transferring pre-trained models to downstream tasks has become a crucial pa…
cs.CL2023
GIST: Improving Parameter Efficient Fine Tuning via Knowledge Interaction
Jiacheng Ruan, Jingsheng Gao, Mingye Xie +4
The Parameter-Efficient Fine-Tuning (PEFT) method, which adjusts or introduces fewer trainable parameters to calibrate pre-trained models on downstream tasks, has become a recent r…
cs.CV2023
Colo-SCRL: Self-Supervised Contrastive Representation Learning for Colonoscopic Video Retrieval
Qingzhong Chen, Shilun Cai, Crystal Cai +3
Colonoscopic video retrieval, which is a critical part of polyp treatment, has great clinical significance for the prevention and treatment of colorectal cancer. However, retrieval…