Facial-Sketch Synthesis: A New Challenge
arXiv:2112.15439 · doi:10.1007/s11633-022-1349-9
Abstract
This paper aims to conduct a comprehensive study on facial-sketch synthesis (FSS). However, due to the high costs of obtaining hand-drawn sketch datasets, there lacks a complete benchmark for assessing the development of FSS algorithms over the last decade. We first introduce a high-quality dataset for FSS, named FS2K, which consists of 2,104 image-sketch pairs spanning three types of sketch styles, image backgrounds, lighting conditions, skin colors, and facial attributes. FS2K differs from previous FSS datasets in difficulty, diversity, and scalability and should thus facilitate the progress of FSS research. Second, we present the largest-scale FSS investigation by reviewing 89 classical methods, including 25 handcrafted feature-based facial-sketch synthesis approaches, 29 general translation methods, and 35 image-to-sketch approaches. Besides, we elaborate comprehensive experiments on the existing 19 cutting-edge models. Third, we present a simple baseline for FSS, named FSGAN. With only two straightforward components, i.e., facial-aware masking and style-vector expansion, FSGAN surpasses the performance of all previous state-of-the-art models on the proposed FS2K dataset by a large margin. Finally, we conclude with lessons learned over the past years and point out several unsolved challenges. Our code is available at https://github.com/DengPingFan/FSGAN.
Accepted to Machine Intelligence Research (MIR)
References in corpus (16)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Conditional Generative Adversarial Nets
- Learning Face Representation from Scratch
- MLP-Mixer: An all-MLP Architecture for Vision
- A Learned Representation For Artistic Style
- A Neural Representation of Sketch Drawings
- End-to-End Photo-Sketch Generation via Fully Convolutional Representation Learning
- Towards the Automatic Anime Characters Creation with Generative Adversarial Networks
- Identity-Aware CycleGAN for Face Photo-Sketch Synthesis and Recognition
- ViTGAN: Training GANs with Vision Transformers
- SofGAN: A Portrait Image Generator with Dynamic Styling
- Comparison and Analysis of Image-to-Image Generative Adversarial Networks: A Survey
- Interactive Sketch & Fill: Multiclass Sketch-to-Image Translation
- PI-REC: Progressive Image Reconstruction Network With Edge and Color Domain
- DeepI2I: Enabling Deep Hierarchical Image-to-Image Translation by Transferring from GANs
- Domain-Specific Mappings for Generative Adversarial Style Transfer
Cited by in corpus (4)
- A survey of synthetic data augmentation methods in computer vision
- Stylized Face Sketch Extraction via Generative Prior with Limited Data
- CLIP4Sketch: Enhancing Sketch to Mugshot Matching through Dataset Augmentation using Diffusion Models
- MixSA: Training-free Reference-based Sketch Extraction via Mixture-of-Self-Attention