3 papers
cs.CV2025
AI killed the video star. Audio-driven diffusion model for expressive talking head generation
Baptiste Chopin, Tashvik Dhamija, Pranav Balaji +2
We propose Dimitra++, a novel framework for audio-driven talking head generation, streamlined to learn lip motion, facial expression, as well as head pose motion. Specifically, we…
cs.CV2025
Do You See What I Say? Generalizable Deepfake Detection based on Visual Speech Recognition
Maheswar Bora, Tashvik Dhamija, Shukesh Reddy +4
Deepfake generation has witnessed remarkable progress, contributing to highly realistic generated images, videos, and audio. While technically intriguing, such progress has raised…
cs.CV2025
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation
Baptiste Chopin, Tashvik Dhamija, Pranav Balaji +2
We propose Dimitra, a novel framework for audio-driven talking head generation, streamlined to learn lip motion, facial expression, as well as head pose motion. Specifically, we tr…