3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CV2026
SVAgent: Storyline-Guided Long Video Understanding via Cross-Modal Multi-Agent Collaboration
Zhongyu Yang, Zuhao Yang, Shuo Zhan +3
Video question answering (VideoQA) is a challenging task that requires integrating spatial, temporal, and semantic information to capture the complex dynamics of video sequences. A…
cs.CL2023★ 3 cited
FACE: Evaluating Natural Language Generation with Fourier Analysis of Cross-Entropy
Zuhao Yang, Yingfang Yuan, Yang Xu +3
Measuring the distance between machine-produced and human language is a critical open problem. Inspired by empirical findings from psycholinguistics on the periodicity of entropy i…