2 citations · 6 across the 5 of their papers we have counts for
1 paper · 1 filter
Zheng Li, Caili Guo, Xin Wang +3
There are two popular loss functions used for vision-language retrieval, i.e., triplet loss and contrastive learning loss, both of them essentially minimize the difference between…