2 papers
cs.CV2025
Global2Local: A Joint-Hierarchical Attention for Video Captioning
Chengpeng Dai, Fuhai Chen, Xiaoshuai Sun +3
Recently, automatic video captioning has attracted increasing attention, where the core challenge lies in capturing the key semantic items, like objects and actions as well as thei…
cs.CV2025
Towards General Visual-Linguistic Face Forgery Detection(V2)
Ke Sun, Shen Chen, Taiping Yao +5
Face manipulation techniques have achieved significant advances, presenting serious challenges to security and social trust. Recent works demonstrate that leveraging multimodal mod…