1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
Just Noticeable Visual Redundancy Forecasting: A Deep Multimodal-driven Approach
Wuyuan Xie, Shukang Wang, Sukun Tian +3
Just noticeable difference (JND) refers to the maximum visual change that human eyes cannot perceive, and it has a wide range of applications in multimedia systems. However, most e…
cs.CV2022
Grafting Pre-trained Models for Multimodal Headline Generation
Lingfeng Qiao, Chen Wu, Ye Liu +3
Multimodal headline utilizes both video frames and transcripts to generate the natural language title of the videos. Due to a lack of large-scale, manually annotated data, the task…
cs.CV2022
OS-MSL: One Stage Multimodal Sequential Link Framework for Scene Segmentation and Classification
Ye Liu, Lingfeng Qiao, Di Yin +4
Scene segmentation and classification (SSC) serve as a critical step towards the field of video structuring analysis. Intuitively, jointly learning of these two tasks can promote e…