activity
20182022
most citedMake-A-Video: Text-to-Video Generation without Text-Video Data

315 citations · 346 across the 5 of their papers we have counts for

collaborators

11 papers

cs.SD20221 cited

Audio Language Modeling using Perceptually-Guided Discrete Representations

Felix Kreuk, Yaniv Taigman, Adam Polyak +4

In this work, we study the task of Audio Language Modeling, in which we aim at learning probabilistic models for audio that can be used for generation and completion. We use a stat…

cs.CV2022315 cited

Make-A-Video: Text-to-Video Generation without Text-Video Data

Uriel Singer, Adam Polyak, Thomas Hayes +10

We propose Make-A-Video -- an approach for directly translating the tremendous recent progress in Text-to-Image (T2I) generation to Text-to-Video (T2V). Our intuition is simple: le…

cs.CV202216 cited

Make-A-Scene: Scene-Based Text-to-Image Generation with Human Priors

Oran Gafni, Adam Polyak, Oron Ashual +3

Recent text-to-image generation methods provide a simple yet exciting conversion capability between text and image domains. While these methods have incrementally improved the gene…

cs.SD2021

High Fidelity Speech Regeneration with Application to Speech Enhancement

Adam Polyak, Lior Wolf, Yossi Adi +2

Speech enhancement has seen great improvement in recent years mainly through contributions in denoising, speaker separation, and dereverberation methods that mostly deal with envir…

eess.AS2020

Unsupervised Cross-Domain Singing Voice Conversion

Adam Polyak, Lior Wolf, Yossi Adi +1

We present a wav-to-wav generative model for the task of singing voice conversion from any identity. Our method utilizes both an acoustic model, trained for the task of automatic s…

cs.LG20194 cited

Live Face De-Identification in Video

Oran Gafni, Lior Wolf, Yaniv Taigman

We propose a method for face de-identification that enables fully automatic video modification at high frame rates. The goal is to maximally decorrelate the identity, while having…