164 citations · 203 across the 21 of their papers we have counts for
29 papers
Learning to Taste: A Multimodal Wine Dataset
Thoranna Bender, Simon Moe Sørensen, Alireza Kashani +5
We present WineSensed, a large multimodal wine dataset for studying the relations between visual perception, language, and flavor. The dataset encompasses 897k images of wine label…
Fashionpedia-Ads: Do Your Favorite Advertisements Reveal Your Fashion Taste?
Mengyun Shi, Claire Cardie, Serge Belongie
Consumers are exposed to advertisements across many different domains on the internet, such as fashion, beauty, car, food, and others. On the other hand, fashion represents second…
Fashionpedia-Taste: A Dataset towards Explaining Human Fashion Taste
Mengyun Shi, Serge Belongie, Claire Cardie
Existing fashion datasets do not consider the multi-facts that cause a consumer to like or dislike a fashion image. Even two consumers like a same fashion image, they could like th…
Discriminative Class Tokens for Text-to-Image Diffusion Models
Idan Schwartz, Vésteinn Snæbjarnarson, Hila Chefer +4
Recent advances in text-to-image diffusion models have enabled the generation of diverse and high-quality images. While impressive, the images often fall short of depicting subtle…
Re-evaluating the Need for Multimodal Signals in Unsupervised Grammar Induction
Boyi Li, Rodolfo Corona, Karttikeya Mangalam +7
Are multimodal inputs necessary for grammar induction? Recent work has shown that multimodal training inputs can improve grammar induction. However, these improvements are based on…
PyTorch Adapt
Kevin Musgrave, Serge Belongie, Ser-Nam Lim
PyTorch Adapt is a library for domain adaptation, a type of machine learning algorithm that re-purposes existing models to work in new domains. It is a fully-featured toolkit, allo…