papers

Publications (7)

cs.CV2022

Cross-Modal Coherence for Text-to-Image Retrieval

Malihe Alikhani, Fangda Han, Hareesh Ravi +3

Common image-text joint understanding techniques presume that images and the associated text can universally be characterized by a single implicit model. However, co-occurring imag…

cs.CV2019

The Art of Food: Meal Image Synthesis from Ingredients

Fangda Han, Ricardo Guerrero, Vladimir Pavlovic

In this work we propose a new computational framework, based on generative deep models, for synthesis of photo-realistic food meal images from textual descriptions of its ingredien…

cs.CV2021

MPG: A Multi-ingredient Pizza Image Generator with Conditional StyleGANs

Fangda Han, Guoyao Hao, Ricardo Guerrero +1

Multilabel conditional image generation is a challenging problem in computer vision. In this work we propose Multi-ingredient Pizza Generator (MPG), a conditional Generative Neural…

cs.CV2021

Multi-attribute Pizza Generator: Cross-domain Attribute Control with Conditional StyleGAN

Fangda Han, Guoyao Hao, Ricardo Guerrero +1

Multi-attribute conditional image generation is a challenging problem in computervision. We propose Multi-attribute Pizza Generator (MPG), a conditional Generative Neural Network (…

cs.CV2020

CookGAN: Meal Image Synthesis from Ingredients

Fangda Han, Ricardo Guerrero, Vladimir Pavlovic

In this work we propose a new computational framework, based on generative deep models, for synthesis of photo-realistic food meal images from textual list of its ingredients. Prev…

cs.CV2019

Cartoonish sketch-based face editing in videos using identity deformation transfer

Long Zhao, Fangda Han, Xi Peng +4

We address the problem of using hand-drawn sketches to create exaggerated deformations to faces in videos, such as enlarging the shape or modifying the position of eyes or mouth. T…