1 paper · 1 filter
Hasan Abed Al Kader Hammoud, Bernard Ghanem
We propose DiffCLIP, a novel vision-language model that extends the differential attention mechanism to CLIP architectures. Differential attention was originally developed for larg…