1 paper
Junze Wang, Lei Fan, Dezheng Zhang +5
Visual Prompt Tuning (VPT) adapts a frozen Vision Transformer (ViT) to downstream tasks by inserting a small number of learnable prompt tokens into the token sequence at each layer…