Publications (7)
Cross-Modal Coherence for Text-to-Image Retrieval
Malihe Alikhani, Fangda Han, Hareesh Ravi +3
Common image-text joint understanding techniques presume that images and the associated text can universally be characterized by a single implicit model. However, co-occurring imag…
The Art of Food: Meal Image Synthesis from Ingredients
Fangda Han, Ricardo Guerrero, Vladimir Pavlovic
In this work we propose a new computational framework, based on generative deep models, for synthesis of photo-realistic food meal images from textual descriptions of its ingredien…
MPG: A Multi-ingredient Pizza Image Generator with Conditional StyleGANs
Fangda Han, Guoyao Hao, Ricardo Guerrero +1
Multilabel conditional image generation is a challenging problem in computer vision. In this work we propose Multi-ingredient Pizza Generator (MPG), a conditional Generative Neural…
Multi-attribute Pizza Generator: Cross-domain Attribute Control with Conditional StyleGAN
Fangda Han, Guoyao Hao, Ricardo Guerrero +1
Multi-attribute conditional image generation is a challenging problem in computervision. We propose Multi-attribute Pizza Generator (MPG), a conditional Generative Neural Network (…
CookGAN: Meal Image Synthesis from Ingredients
Fangda Han, Ricardo Guerrero, Vladimir Pavlovic
In this work we propose a new computational framework, based on generative deep models, for synthesis of photo-realistic food meal images from textual list of its ingredients. Prev…
Cartoonish sketch-based face editing in videos using identity deformation transfer
Long Zhao, Fangda Han, Xi Peng +4
We address the problem of using hand-drawn sketches to create exaggerated deformations to faces in videos, such as enlarging the shape or modifying the position of eyes or mouth. T…