Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
PreMind: Multi-Agent Video Understanding for Advanced Indexing of Presentation-style Videos
Kangda Wei, Zhengyu Zhou, Bingqing Wang +4
In recent years, online lecture videos have become an increasingly popular resource for acquiring new knowledge. Systems capable of effectively understanding/indexing lecture video…
cs.CV2024
UAL-Bench: The First Comprehensive Unusual Activity Localization Benchmark
Hasnat Md Abdullah, Tian Liu, Kangda Wei +2
Localizing unusual activities, such as human errors or surveillance incidents, in videos holds practical significance. However, current video understanding models struggle with loc…