10 citations · 12 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 10 cited
Efficient-VDVAE: Less is more
Louay Hazami, Rayhane Mama, Ragavan Thurairatnam
Hierarchical VAEs have emerged in recent years as a reliable option for maximum likelihood estimation. However, instability issues and demanding computational requirements have hin…
cs.SD2021★ 2 cited
NWT: Towards natural audio-to-video generation with representation learning
Rayhane Mama, Marc S. Tyndel, Hashiam Kadhim +2
In this work we introduce NWT, an expressive speech-to-video model. Unlike approaches that use domain-specific intermediate representations such as pose keypoints, NWT learns its o…