Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
LeVLJEPA: End-to-End Vision-Language Pretraining Without Negatives
Lukas Kuhn, Giuseppe Serra, Randall Balestriero +1
Vision-language pretraining remains dominated by contrastive objectives, whereas vision-only self-supervised learning has largely adopted non-contrastive methods. At the same time,…
cs.CV2026
Non-Contrastive Vision-Language Learning with Predictive Embedding Alignment
Lukas Kuhn, Giuseppe Serra, Florian Buettner
Vision-language models have transformed multimodal representation learning, yet dominant contrastive approaches like CLIP require large batch sizes, careful negative sampling, and…