8 papers
Audio-Based Understanding of Audiobook Narration Appeal
Shahar Elisha, Mariano Beguerisse-Díaz, Emmanouil Benetos
Narration is central to the audiobook listening experience, shaping how listeners engage with and understand the content. This work explores how narration qualities shape an audiob…
ViSMaP: Unsupervised Hour-long Video Summarisation by Meta-Prompting
Jian Hu, Dimitrios Korkinof, Shaogang Gong +1
We introduce ViSMap: Unsupervised Video Summarisation by Meta Prompting, a system to summarise hour long videos with no-supervision. Most existing video understanding models work w…
Classification of Spontaneous and Scripted Speech for Multilingual Audio
Shahar Elisha, Andrew McDowell, Mariano Beguerisse-Díaz +1
Distinguishing scripted from spontaneous speech is an essential tool for better understanding how speech styles influence speech processing research. It can also improve recommenda…
Link Me Baby One More Time: Social Music Discovery on Spotify
Shazia'Ayn Babul, Desislava Hristova, Antonio Lima +2
We explore the social and contextual factors that influence the outcome of person-to-person music recommendations and discovery. Specifically, we use data from Spotify to investiga…
Topological fingerprints for audio identification
Wojciech Reise, Ximena Fernández, Maria Dominguez +2
We present a topological audio fingerprinting approach for robustly identifying duplicate audio tracks. Our method applies persistent homology on local spectral decompositions of a…
Thematic recommendations on knowledge graphs using multilayer networks
Mariano Beguerisse-Díaz, Dimitrios Korkinof, Till Hoffmann
We present a framework to generate and evaluate thematic recommendations based on multilayer network representations of knowledge graphs (KGs). In this representation, each layer e…