1 paper
Mohamed Fazli Imam, Rufael Fedaku Marew, Jameel Hassan +3
In the era of foundation models, CLIP has emerged as a powerful tool for aligning text & visual modalities into a common embedding space. However, the alignment objective used to t…