activity
20212025
most citedLeveraging Hierarchical Structures for Few-Shot Musical Instrument Recognition

7 citations · 17 across the 5 of their papers we have counts for

collaborators

7 papers

cs.LG2025

WhAM: Towards A Translative Model of Sperm Whale Vocalization

Orr Paradise, Pranav Muralikrishnan, Liangyuan Chen +6

Sperm whales communicate in short sequences of clicks known as codas. We present WhAM (Whale Acoustics Model), the first transformer-based model capable of generating synthetic spe…

eess.AS2025

HARP 2.0: Expanding Hosted, Asynchronous, Remote Processing for Deep Learning in the DAW

Christodoulos Benetatos, Frank Cwitkowitz, Nathan Pruyne +4

HARP 2.0 brings deep learning models to digital audio workstation (DAW) software through hosted, asynchronous, remote processing, allowing users to route audio from a plug-in inter…

cs.SD2024

Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations

Hugo Flores García, Oriol Nieto, Justin Salamon +2

We present Sketch2Sound, a generative audio model capable of creating high-quality sounds from a set of interpretable time-varying control signals: loudness, brightness, and pitch,…

cs.SD2024★ 3 cited

Exploring Musical Roots: Applying Audio Embeddings to Empower Influence Attribution for a Generative Music Model

Julia Barnett, Hugo Flores Garcia, Bryan Pardo

Every artist has a creative process that draws inspiration from previous artists and their works. Today, "inspiration" has been automated by generative music models. The black box…

cs.SD2023★ 7 cited

VampNet: Music Generation via Masked Acoustic Token Modeling

Hugo Flores Garcia, Prem Seetharaman, Rithesh Kumar +1

We introduce VampNet, a masked acoustic token modeling approach to music synthesis, compression, inpainting, and variation. We use a variable masking schedule during training which…

cs.SD2021

Deep Learning Tools for Audacity: Helping Researchers Expand the Artist's Toolkit

Hugo Flores Garcia, Aldo Aguilar, Ethan Manilow +2

We present a software framework that integrates neural networks into the popular open-source audio editing software, Audacity, with a minimal amount of developer effort. In this pa…