3 papers
cs.CV2026
The TIME Machine: On The Power of Motion for Efficient Perception
Mantas Skackauskas, Xinyue Hao, Laura Sevilla-Lara
Video representation learning has seen tremendous progress in recent years. This has been driven by many factors, including the scale of training and the success of visual models t…
cs.CV2026
Why Do Vision Language Models Struggle To Recognize Human Emotions?
Madhav Agarwal, Sotirios A. Tsaftaris, Laura Sevilla-Lara +1
Understanding emotions is a fundamental ability for intelligent systems to be able to interact with humans. Vision-language models (VLMs) have made tremendous progress in the last…
cs.CV2026
It's a Matter of Time: Three Lessons on Long-Term Motion for Perception
Willem Davison, Xinyue Hao, Laura Sevilla-Lara
Temporal information has long been considered to be essential for perception. While there is extensive research on the role of image information for perceptual tasks, the role of t…