Face Generation and Editing with StyleGAN: A Survey
arXiv:2212.09102 · doi:10.1109/TPAMI.2024.3350004
Abstract
Our goal with this survey is to provide an overview of the state of the art deep learning methods for face generation and editing using StyleGAN. The survey covers the evolution of StyleGAN, from PGGAN to StyleGAN3, and explores relevant topics such as suitable metrics for training, different latent representations, GAN inversion to latent spaces of StyleGAN, face image editing, cross-domain face stylization, face restoration, and even Deepfake applications. We aim to provide an entry point into the field for readers that have basic knowledge about the field of deep learning and are looking for an accessible introduction and overview.
References in corpus (17)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Hierarchical Text-Conditional Image Generation with CLIP Latents
- Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
- Alias-Free Generative Adversarial Networks
- Blended Latent Diffusion
- Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold
- StyleNeRF: A Style-based 3D-Aware Generator for High-resolution Image Synthesis
- LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
- Collaborative Learning for Faster StyleGAN Embedding
- StyleGAN-T: Unlocking the Power of GANs for Fast Large-Scale Text-to-Image Synthesis
- Resolution Dependent GAN Interpolation for Controllable Image Synthesis Between Domains
- StyleCariGAN: Caricature Generation via StyleGAN Feature Map Modulation
- Diffused Heads: Diffusion Models Beat GANs on Talking-Face Generation
- Unconstrained Facial Expression Transfer using Style-based Generator
- 3D Cartoon Face Generation with Controllable Expressions from a Single GAN Image
- MyStyle: A Personalized Generative Prior
- Faces: AI Blitz XIII Solutions