3 papers
cs.CV2025
ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion
Hoang-Son Vo, Quang-Vinh Nguyen, Seungwon Kim +3
Audio-driven talking head generation requires precise synchronization between facial animations and audio signals. This paper introduces ATL-Diff, a novel approach addressing synch…
cs.LG2025
Latent Behavior Diffusion for Sequential Reaction Generation in Dyadic Setting
Minh-Duc Nguyen, Hyung-Jeong Yang, Soo-Hyung Kim +2
The dyadic reaction generation task involves synthesizing responsive facial reactions that align closely with the behaviors of a conversational partner, enhancing the naturalness a…
cs.CV2024
Transformer with Leveraged Masked Autoencoder for video-based Pain Assessment
Minh-Duc Nguyen, Hyung-Jeong Yang, Soo-Hyung Kim +2
Accurate pain assessment is crucial in healthcare for effective diagnosis and treatment; however, traditional methods relying on self-reporting are inadequate for populations unabl…