3 papers
cs.CV2025
ChineseVideoBench: Benchmarking Multi-modal Large Models for Chinese Video Question Answering
Yuxiang Nie, Han Wang, Yongjie Ye +15
This paper introduces ChineseVideoBench, a pioneering benchmark specifically designed for evaluating Multimodal Large Language Models (MLLMs) in Chinese Video Question Answering. T…
cs.IR2025
SUMMA: A Multimodal Large Language Model for Advertisement Summarization
Weitao Jia, Shuo Yin, Zhoufutu Wen +6
Understanding multimodal video ads is crucial for improving query-ad matching and relevance ranking on short video platforms, enhancing advertising effectiveness and user experienc…
cs.CL2024
ChatASU: Evoking LLM's Reflexion to Truly Understand Aspect Sentiment in Dialogues
Yiding Liu, Jingjing Wang, Jiamin Luo +2
Aspect Sentiment Understanding (ASU) in interactive scenarios (e.g., Question-Answering and Dialogue) has attracted ever-more interest in recent years and achieved important progre…