3 papers
cs.CV2025
MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control
Fatemeh Nazarieh, Zhenhua Feng, Diptesh Kanojia +2
Audio-driven talking face generation has gained significant attention for applications in digital media and virtual avatars. While recent methods improve audio-lip synchronization,…
cs.CL2025
The Mind's Eye: A Multi-Faceted Reward Framework for Guiding Visual Metaphor Generation
Girish A. Koushik, Fatemeh Nazarieh, Katherine Birch +2
Visual metaphor generation is a challenging task that aims to generate an image given an input text metaphor. Inherently, it needs language understanding to bind a source concept w…
cs.CV2024
PortraitTalk: Towards Customizable One-Shot Audio-to-Talking Face Generation
Fatemeh Nazarieh, Zhenhua Feng, Diptesh Kanojia +2
Audio-driven talking face generation is a challenging task in digital communication. Despite significant progress in the area, most existing methods concentrate on audio-lip synchr…