Collaging Class-specific GANs for Semantic Image Synthesis
arXiv:2110.04281
Abstract
We propose a new approach for high resolution semantic image synthesis. It consists of one base image generator and multiple class-specific generators. The base generator generates high quality images based on a segmentation map. To further improve the quality of different objects, we create a bank of Generative Adversarial Networks (GANs) by separately training class-specific models. This has several benefits including -- dedicated weights for each class; centrally aligned data for each model; additional training data from other sources, potential of higher resolution and quality; and easy manipulation of a specific object in the scene. Experiments show that our approach can generate high quality images in high resolution while having flexibility of object-level control by using class-specific generators.
ICCV 2021
References in corpus (7)
- GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
- Class-Balanced Loss Based on Effective Number of Samples
- Learning to Predict Layout-to-image Conditional Convolutions for Semantic Image Synthesis
- You Only Need Adversarial Supervision for Semantic Image Synthesis
- Style and Pose Control for Image Synthesis of Humans from a Single Monocular View
- Global and Local Consistent Age Generative Adversarial Networks
- Mask-Guided Portrait Editing with Conditional GANs