3 citations · 5 across the 2 of their papers we have counts for
7 papers
Towards Multi-domain Face Landmark Detection with Synthetic Data from Diffusion model
Yuanming Li, Gwantae Kim, Jeong-gi Kwak +2
Recently, deep learning-based facial landmark detection for in-the-wild faces has achieved significant improvement. However, there are still challenges in face landmark detection i…
MPE4G: Multimodal Pretrained Encoder for Co-Speech Gesture Generation
Gwantae Kim, Seonghyeok Noh, Insung Ham +1
When virtual agents interact with humans, gestures are crucial to delivering their intentions with speech. Previous multimodal co-speech gesture generation models required encoded…
DIFAI: Diverse Facial Inpainting using StyleGAN Inversion
Dongsik Yoon, Jeong-gi Kwak, Yuanming Li +2
Image inpainting is an old problem in computer vision that restores occluded regions and completes damaged images. In the case of facial image inpainting, most of the methods gener…
Reference Guided Image Inpainting using Facial Attributes
Dongsik Yoon, Jeonggi Kwak, Yuanming Li +3
Image inpainting is a technique of completing missing pixels such as occluded region restoration, distracting objects removal, and facial completion. Among these inpainting tasks,…
Injecting 3D Perception of Controllable NeRF-GAN into StyleGAN for Editable Portrait Image Synthesis
Jeong-gi Kwak, Yuanming Li, Dongsik Yoon +3
Over the years, 2D GANs have achieved great successes in photorealistic portrait generation. However, they lack 3D understanding in the generation process, thus they suffer from mu…
Non-negative matrix factorization-based subband decomposition for acoustic source localization
Suwon Shon, Seongkyu Mun, David Han +1
A novel non-negative matrix factorization (NMF) based subband decomposition in frequency spatial domain for acoustic source localization using a microphone array is introduced. The…