14 citations · 14 across the 2 of their papers we have counts for
2 papers
cs.SD2024
MusicFlow: Cascaded Flow Matching for Text Guided Music Generation
K R Prajwal, Bowen Shi, Matthew Lee +8
We introduce MusicFlow, a cascaded text-to-music generation model based on flow matching. Based on self-supervised representations to bridge between text descriptions and music aud…
cs.CV2022★ 14 cited
Lip-to-Speech Synthesis for Arbitrary Speakers in the Wild
Sindhu B Hegde, K R Prajwal, Rudrabha Mukhopadhyay +2
In this work, we address the problem of generating speech from silent lip videos for any speaker in the wild. In stark contrast to previous works, our method (i) is not restricted…