3 papers
cs.CV2025
REVEAL: Relation-based Video Representation Learning for Video-Question-Answering
Sofian Chaybouti, Walid Bousselham, Moritz Wolter +1
Video-Question-Answering (VideoQA) comprises the capturing of complex visual relation changes over time, remaining a challenge even for advanced Video Language Models (VLM), i.a.,…
cs.SE2025
More Rigorous Software Engineering Would Improve Reproducibility in Machine Learning Research
Moritz Wolter, Lokesh Veeramacheneni, Charles Tapley Hoyt
While experimental reproduction remains a pillar of the scientific method, we observe that the software best practices supporting the reproduction of machine learning ( ML ) resear…
cs.CV2024
Rethinking temporal self-similarity for repetitive action counting
Yanan Luo, Jinhui Yi, Yazan Abu Farha +2
Counting repetitive actions in long untrimmed videos is a challenging task that has many applications such as rehabilitation. State-of-the-art methods predict action counts by firs…