568 citations
- Institució Catalana de Recerca i Estudis AvançatsES41 papers
- Santa Fe InstituteUS15 papers
- Universitat Politècnica de CatalunyaES11 papers
- Barcelona Biomedical Research ParkES10 papers
- Institut de Biologia EvolutivaES7 papers
- Centre National de la Recherche ScientifiqueFR6 papers
- Universitat de BarcelonaES6 papers
- Centre de Recerca MatemàticaES5 papers
- University of OxfordGB5 papers
- Institut national de recherche en sciences et technologies du numériqueFR4 papers
- McGill UniversityCA4 papers
- New York UniversityUS4 papers
17 papers · 1 filter
Efficient and Fast Generative-Based Singing Voice Separation using a Latent Diffusion Model
Genís Plaja-Roglans, Yun-Ning Hung, Xavier Serra +1
Extracting individual elements from music mixtures is a valuable tool for music production and practice. While neural networks optimized to mask or transform mixture spectrograms i…
Music Rearrangement Using Hierarchical Segmentation
Christos Plachouras, Marius Miron
Music rearrangement involves reshuffling, deleting, and repeating sections of a music piece with the goal of producing a standalone version that has a different duration. It is a c…
MusicLM: Generating Music From Text
Andrea Agostinelli, Timo I. Denk, Zalán Borsos +10
We introduce MusicLM, a model generating high-fidelity music from text descriptions such as "a calming violin melody backed by a distorted guitar riff". MusicLM casts the process o…
Audio-based Musical Version Identification: Elements and Challenges
Furkan Yesiler, Guillaume Doras, Rachel M. Bittner +2
In this article, we aim to provide a review of the key ideas and approaches proposed in 20 years of scientific literature around musical version identification (VI) research and co…
Improving Sound Event Classification by Increasing Shift Invariance in Convolutional Neural Networks
Eduardo Fonseca, Andres Ferraro, Xavier Serra
Recent studies have put into question the commonly assumed shift invariance property of convolutional networks, showing that small shifts in the input can affect the output predict…
Enriched Music Representations with Multiple Cross-modal Contrastive Learning
Andres Ferraro, Xavier Favory, Konstantinos Drossos +2
Modeling various aspects that make a music piece unique is a challenging task, requiring the combination of multiple sources of information. Deep learning is commonly used to obtai…