32 citations · 78 across the 9 of their papers we have counts for
3 papers · 1 filter
Zero-Shot Dense Video Captioning by Jointly Optimizing Text and Moment
Yongrae Jo, Seongyun Lee, Aiden SJ Lee +3
Dense video captioning, a task of localizing meaningful moments and generating relevant captions for videos, often requires a large, expensive corpus of annotated video segments pa…
ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation
Aiden Seungjoon Lee, Hanseok Oh, Minjoon Seo
Video-text retrieval has many real-world applications such as media analytics, surveillance, and robotics. This paper presents the 1st place solution to the video retrieval track o…
A Diagram Is Worth A Dozen Images
Aniruddha Kembhavi, Mike Salvato, Eric Kolve +3
Diagrams are common tools for representing complex concepts, relationships and events, often when it would be difficult to portray the same information with natural images. Underst…