14 citations · 47 across the 13 of their papers we have counts for
5 papers · 1 filter
Deep Learning Tools for Audacity: Helping Researchers Expand the Artist's Toolkit
Hugo Flores Garcia, Aldo Aguilar, Ethan Manilow +2
We present a software framework that integrates neural networks into the popular open-source audio editing software, Audacity, with a minimal amount of developer effort. In this pa…
Unsupervised Source Separation By Steering Pretrained Music Models
Ethan Manilow, Patrick O'Reilly, Prem Seetharaman +1
We showcase an unsupervised method that repurposes deep models trained for music generation and music tagging for audio source separation, without any retraining. An audio generati…
Neural Pitch-Shifting and Time-Stretching with Controllable LPCNet
Max Morrison, Zeyu Jin, Nicholas J. Bryan +2
Modifying the pitch and timing of an audio signal are fundamental audio editing operations with applications in speech manipulation, audio-visual synchronization, and singing voice…
Leveraging Hierarchical Structures for Few-Shot Musical Instrument Recognition
Hugo Flores Garcia, Aldo Aguilar, Ethan Manilow +1
Deep learning work on musical instrument recognition has generally focused on instrument classes for which we have abundant data. In this work, we exploit hierarchical relationship…
Context-Aware Prosody Correction for Text-Based Speech Editing
Max Morrison, Lucas Rencker, Zeyu Jin +3
Text-based speech editors expedite the process of editing speech recordings by permitting editing via intuitive cut, copy, and paste operations on a speech transcript. A major draw…