36 citations · 43 across the 5 of their papers we have counts for
4 papers · 1 filter
HEAR: Holistic Evaluation of Audio Representations
Joseph Turian, Jordie Shier, Humair Raj Khan +20
What audio embedding approach generalizes best to a wide range of downstream tasks across a variety of everyday domains without fine-tuning? The aim of the HEAR benchmark is to dev…
One Billion Audio Sounds from GPU-enabled Modular Synthesis
Joseph Turian, Jordie Shier, George Tzanetakis +2
We release synth1B1, a multi-modal audio corpus consisting of 1 billion 4-second synthesized sounds, paired with the synthesis parameters used to generate them. The dataset is 100x…
Deep Autotuner: a Pitch Correcting Network for Singing Performances
Sanna Wager, George Tzanetakis, Cheng-i Wang +1
We introduce a data-driven approach to automatic pitch correction of solo singing performances. The proposed approach predicts note-wise pitch shifts from the relationship between…
Deep Autotuner: A Data-Driven Approach to Natural-Sounding Pitch Correction for Singing Voice in Karaoke Performances
Sanna Wager, George Tzanetakis, Cheng-i Wang +3
We describe a machine-learning approach to pitch correcting a solo singing performance in a karaoke setting, where the solo voice and accompaniment are on separate tracks. The prop…