38 citations · 43 across the 3 of their papers we have counts for
4 papers
Establishing an Evaluation Metric to Quantify Climate Change Image Realism
Sharon Zhou, Alexandra Luccioni, Gautier Cosne +2
With success on controlled tasks, generative models are being increasingly applied to humanitarian applications [1,2]. In this paper, we focus on the evaluation of a conditional ge…
Boomerang: Rebounding the Consequences of Reputation Feedback on Crowdsourcing Platforms
Snehalkumar, S. Gaikwad, Durim Morina +35
Paid crowdsourcing platforms suffer from low-quality work and unfair rejections, but paradoxically, most workers and requesters have high reputation scores. These inflated scores,…
HYPE: A Benchmark for Human eYe Perceptual Evaluation of Generative Models
Sharon Zhou, Mitchell L. Gordon, Ranjay Krishna +3
Generative models often use human evaluations to measure the perceived quality of their outputs. Automated metrics are noisy indirect proxies, because they rely on heuristics or pr…
Prototype Tasks: Improving Crowdsourcing Results through Rapid, Iterative Task Design
Snehalkumar "Neil" S. Gaikwad, Nalin Chhibber, Vibhor Sehgal +26
Low-quality results have been a long-standing problem on microtask crowdsourcing platforms, driving away requesters and justifying low wages for workers. To date, workers have been…