3 papers
cs.CV2026
Breaking the Limits of Open-Weight CLIP: An Optimization Framework for Self-supervised Fine-tuning of CLIP
Anant Mehta, Xiyuan Wei, Xingyu Chen +1
CLIP has become a cornerstone of multimodal representation learning, yet improving its performance typically requires a prohibitively costly process of training from scratch on bil…
cs.CV2025
HFMF: Hierarchical Fusion Meets Multi-Stream Models for Deepfake Detection
Anant Mehta, Bryant McArthur, Nagarjuna Kolloju +1
The rapid progress in deep generative models has led to the creation of incredibly realistic synthetic images that are becoming increasingly difficult to distinguish from real-worl…
cs.LG2024
AmCLR: Unified Augmented Learning for Cross-Modal Representations
Ajay Jagannath, Aayush Upadhyay, Anant Mehta
Contrastive learning has emerged as a pivotal framework for representation learning, underpinning advances in both unimodal and bimodal applications like SimCLR and CLIP. To addres…