7 citations · 12 across the 2 of their papers we have counts for
2 papers
cs.AI2025★ 5 cited
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Mido Assran, Adrien Bardes, David Fan +27
A major challenge for modern AI is to learn to understand the world and learn to act largely by observation. This paper explores a self-supervised approach that combines internet-s…
cs.CV2024★ 7 cited
Learning and Leveraging World Models in Visual Representation Learning
Quentin Garrido, Mahmoud Assran, Nicolas Ballas +3
Joint-Embedding Predictive Architecture (JEPA) has emerged as a promising self-supervised approach that learns by leveraging a world model. While previously limited to predicting m…