5 citations · 5 across the 1 of their papers we have counts for
3 papers
cs.CV2026
Not All Inputs Are Valid: Towards Open-Set Video Moment Retrieval Using Language
Xiang Fang, Wanlong Fang, Daizong Liu +8
Video Moment Retrieval (VMR) targets to retrieve the specific moment corresponding to a sentence query from an untrimmed video. Although recent works have made remarkable progress…
cs.CV2026
Fewer Steps, Better Performance: Efficient Cross-Modal Clip Trimming for Video Moment Retrieval Using Language
Xiang Fang, Daizong Liu, Wanlong Fang +5
Given an untrimmed video and a sentence query, video moment retrieval using language (VMR) aims to locate a target query-relevant moment. Since the untrimmed video is overlong, alm…
cs.MM2026★ 5 cited
Hierarchical Local-Global Transformer for Temporal Sentence Grounding
Xiang Fang, Daizong Liu, Pan Zhou +2
This paper studies the multimedia problem of temporal sentence grounding (TSG), which aims to accurately determine the specific video segment in an untrimmed video according to a g…