2 papers
cs.CV2024
Towards Multi-domain Face Landmark Detection with Synthetic Data from Diffusion model
Yuanming Li, Gwantae Kim, Jeong-gi Kwak +2
Recently, deep learning-based facial landmark detection for in-the-wild faces has achieved significant improvement. However, there are still challenges in face landmark detection i…
cs.CV2023
MPE4G: Multimodal Pretrained Encoder for Co-Speech Gesture Generation
Gwantae Kim, Seonghyeok Noh, Insung Ham +1
When virtual agents interact with humans, gestures are crucial to delivering their intentions with speech. Previous multimodal co-speech gesture generation models required encoded…