1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2026
MEDit-Bench: A Dataset for Evaluating Message-Driven Narrative Video Editing
Katsuya Ogata, Zongshang Pang, Mayu Otani +1
Video editing is fundamentally message-driven: even from the same source footage, the selected shots change depending on the narrative the editor wishes to convey. Benchmarks for a…
cs.CV2025
Measure Twice, Cut Once: A Semantic-Oriented Approach to Video Temporal Localization with Video LLMs
Zongshang Pang, Mayu Otani, Yuta Nakashima
Temporally localizing user-queried events through natural language is a crucial capability for video models. Recent methods predominantly adapt video LLMs to generate event boundar…
cs.CV2022★ 1 cited
Contrastive Losses Are Natural Criteria for Unsupervised Video Summarization
Zongshang Pang, Yuta Nakashima, Mayu Otani +1
Video summarization aims to select the most informative subset of frames in a video to facilitate efficient video browsing. Unsupervised methods usually rely on heuristic training…