48 citations · 48 across the 2 of their papers we have counts for
2 papers
cs.CV2024
Exploring Temporal Event Cues for Dense Video Captioning in Cyclic Co-learning
Zhuyang Xie, Yan Yang, Yankai Yu +3
Dense video captioning aims to detect and describe all events in untrimmed videos. This paper presents a dense video captioning network called Multi-Concept Cyclic Learning (MCCL),…
cs.CV2024★ 48 cited
Trustworthy Multimodal Fusion for Sentiment Analysis in Ordinal Sentiment Space
Zhuyang Xie, Yan Yang, Jie Wang +2
Multimodal video sentiment analysis aims to integrate multiple modal information to analyze the opinions and attitudes of speakers. Most previous work focuses on exploring the sema…