Only a Matter of Style: Age Transformation Using a Style-Based Regression Model
arXiv:2102.02754
Abstract
The task of age transformation illustrates the change of an individual's appearance over time. Accurately modeling this complex transformation over an input facial image is extremely challenging as it requires making convincing, possibly large changes to facial features and head shape, while still preserving the input identity. In this work, we present an image-to-image translation method that learns to directly encode real facial images into the latent space of a pre-trained unconditional GAN (e.g., StyleGAN) subject to a given aging shift. We employ a pre-trained age regression network to explicitly guide the encoder in generating the latent codes corresponding to the desired age. In this formulation, our method approaches the continuous aging process as a regression task between the input age and desired target age, providing fine-grained control over the generated image. Moreover, unlike approaches that operate solely in the latent space using a prior on the path controlling age, our method learns a more disentangled, non-linear path. Finally, we demonstrate that the end-to-end nature of our approach, coupled with the rich semantic latent space of StyleGAN, allows for further editing of the generated images. Qualitative and quantitative evaluations show the advantages of our method compared to state-of-the-art approaches.
Accepted to SIGGRAPH 2021, project page available at https://yuval-alaluf.github.io/SAM/
References in corpus (9)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Conditional Generative Adversarial Nets
- On the Variance of the Adaptive Learning Rate and Beyond
- Unsupervised Discovery of Interpretable Directions in the GAN Latent Space
- Age Progression/Regression by Conditional Adversarial Autoencoder
- Few-Shot Unsupervised Image-to-Image Translation
- Semantic Hierarchy Emerges in Deep Generative Representations for Scene Synthesis
- AttentionGAN: Unpaired Image-to-Image Translation using Attention-Guided Generative Adversarial Networks
- StyleGAN2 Distillation for Feed-forward Image Manipulation