1 paper
Yifan Yang, Zhen Zhang, Rupak Vignesh Swaminathan +3
Fine-tuning vision language models (VLMs) has achieved remarkable performance across various downstream tasks; yet, it requires access to model gradients through backpropagation (B…