2 papers
cs.AI2025
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Mido Assran, Adrien Bardes, David Fan +27
A major challenge for modern AI is to learn to understand the world and learn to act largely by observation. This paper explores a self-supervised approach that combines internet-s…
cs.CV2025
CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models
Aaron Foss, Chloe Evans, Sasha Mitts +3
We introduce CausalVQA, a benchmark dataset for video question answering (VQA) composed of question-answer pairs that probe models' understanding of causality in the physical world…