12 papers
Multi-Modal Semantic Expansion with Constrained LLM Reranking for Conversational Music Recommendation
Naman Garg, Sarika Jain, George Fazekas
We present Team Semiintelligencn's solution for the ACM RecSys 2026 TalkPlayData Challenge, addressing conversational music recommendation through a multi-modal and personalized co…
LiveBand: Live Accompaniment Generation in the Audio Domain
Marco Pasini, Javier Nistal, Ben Hayes +3
We present LiveBand, a real-time system that generates high-fidelity music accompaniments to live audio input, respecting strict causal constraints. Our method trains a causal tran…
Diffusion Timbre Transfer Via Mutual Information Guided Inpainting
Ching Ho Lee, Javier Nistal, Stefan Lattner +2
We study timbre transfer as an inference-time editing problem for music audio. Starting from a strong pre-trained latent diffusion model, we introduce a lightweight procedure that…
CoDiCodec: Unifying Continuous and Discrete Compressed Representations of Audio
Marco Pasini, Stefan Lattner, George Fazekas
Efficiently representing audio signals in a compressed latent space is critical for latent generative modelling. However, existing autoencoders often force a choice between continu…
Towards a Unified Representation Evaluation Framework Beyond Downstream Tasks
Christos Plachouras, Julien Guinot, George Fazekas +3
Downstream probing has been the dominant method for evaluating model representations, an important process given the increasing prominence of self-supervised learning and foundatio…
Music2Latent2: Audio Compression with Summary Embeddings and Autoregressive Decoding
Marco Pasini, Stefan Lattner, George Fazekas
Efficiently compressing high-dimensional audio signals into a compact and informative latent space is crucial for various tasks, including generative modeling and music information…