90 citations · 250 across the 28 of their papers we have counts for
5 papers · 1 filter
Zero-Shot Audio Captioning via Audibility Guidance
Tal Shaharabany, Ariel Shaulov, Lior Wolf
The task of audio captioning is similar in essence to tasks such as image and video captioning. However, it has received much less attention. We propose three desiderata for captio…
TTS Skins: Speaker Conversion via ASR
Adam Polyak, Lior Wolf, Yaniv Taigman
We present a fully convolutional wav-to-wav network for converting between speakers' voices, without relying on text. Our network is based on an encoder-decoder architecture, where…
Speech Denoising by Accumulating Per-Frequency Modeling Fluctuations
Michael Michelashvili, Lior Wolf
We present a method for audio denoising that combines processing done in both the time domain and the time-frequency domain. Given a noisy audio clip, the method trains a deep neur…
Semi-Supervised Monaural Singing Voice Separation With a Masking Network Trained on Synthetic Mixtures
Michael Michelashvili, Sagie Benaim, Lior Wolf
We study the problem of semi-supervised singing voice separation, in which the training data contains a set of samples of mixed music (singing and instrumental) and an unmatched se…
A Universal Music Translation Network
Noam Mor, Lior Wolf, Adam Polyak +1
We present a method for translating music across musical instruments, genres, and styles. This method is based on a multi-domain wavenet autoencoder, with a shared encoder and a di…