Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations
Dongqi Liu, Chenxi Whitehouse, Xi Yu +6
Transforming recorded videos into concise and accurate textual summaries is a growing challenge in multimodal learning. This paper introduces VISTA, a dataset specifically designed…
cs.CL2024
A Modular Approach for Multimodal Summarization of TV Shows
Louis Mahon, Mirella Lapata
In this paper we address the task of summarizing television shows, which touches key areas in AI research: complex reasoning, multiple modalities, and long narratives. We present a…