8 citations · 8 across the 1 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025★ 8 cited
Video Understanding with Large Language Models: A Survey
Yolo Y. Tang, Jing Bi, Siting Xu +17
With the burgeoning growth of online video platforms and the escalating volume of video content, the demand for proficient video understanding tools has intensified markedly. Given…
cs.CV2025
Multi-modal Segment Assemblage Network for Ad Video Editing with Importance-Coherence Reward
Yolo Yunlong Tang, Siting Xu, Teng Wang +3
Advertisement video editing aims to automatically edit advertising videos into shorter videos while retaining coherent content and crucial information conveyed by advertisers. It m…