3 papers
cs.CV2026
See More, Store Less: Memory-Efficient Resolution for Video Moment Retrieval
Mingyu Jeon, Sungjin Han, Jinkwon Hwang +3
Recent advances in Multimodal Large Language Models (MLLMs) have improved image recognition and reasoning, but video-related tasks remain challenging due to memory constraints from…
cs.CV2026
GranAlign: Granularity-Aware Alignment Framework for Zero-Shot Video Moment Retrieval
Mingyu Jeon, Sunjae Yoon, Jonghee Kim +1
Zero-shot video moment retrieval (ZVMR) is the task of localizing a temporal moment within an untrimmed video using a natural language query without relying on task-specific traini…
cs.CV2025
Point to Span: Zero-Shot Moment Retrieval for Navigating Unseen Hour-Long Videos
Mingyu Jeon, Jisoo Yang, Sungjin Han +4
Zero-shot Long Video Moment Retrieval (ZLVMR) is the task of identifying temporal segments in hour-long videos using a natural language query without task-specific training. The co…