activity
20242026
collaborators

12 papers

cs.AI2026

Multi-Modal Semantic Expansion with Constrained LLM Reranking for Conversational Music Recommendation

Naman Garg, Sarika Jain, George Fazekas

We present Team Semiintelligencn's solution for the ACM RecSys 2026 TalkPlayData Challenge, addressing conversational music recommendation through a multi-modal and personalized co…

cs.SD2026

LiveBand: Live Accompaniment Generation in the Audio Domain

Marco Pasini, Javier Nistal, Ben Hayes +3

We present LiveBand, a real-time system that generates high-fidelity music accompaniments to live audio input, respecting strict causal constraints. Our method trains a causal tran…

cs.SD2026

Diffusion Timbre Transfer Via Mutual Information Guided Inpainting

Ching Ho Lee, Javier Nistal, Stefan Lattner +2

We study timbre transfer as an inference-time editing problem for music audio. Starting from a strong pre-trained latent diffusion model, we introduce a lightweight procedure that…

cs.SD2025

CoDiCodec: Unifying Continuous and Discrete Compressed Representations of Audio

Marco Pasini, Stefan Lattner, George Fazekas

Efficiently representing audio signals in a compressed latent space is critical for latent generative modelling. However, existing autoencoders often force a choice between continu…

cs.LG2025

Towards a Unified Representation Evaluation Framework Beyond Downstream Tasks

Christos Plachouras, Julien Guinot, George Fazekas +3

Downstream probing has been the dominant method for evaluating model representations, an important process given the increasing prominence of self-supervised learning and foundatio…

cs.SD2025

Music2Latent2: Audio Compression with Summary Embeddings and Autoregressive Decoding

Marco Pasini, Stefan Lattner, George Fazekas

Efficiently compressing high-dimensional audio signals into a compact and informative latent space is crucial for various tasks, including generative modeling and music information…