1 paper
Till Aczel, Lucas Theis, Roger Wattenhofer
Evaluating generative models is challenging because standard metrics often fail to reflect human preferences. Human evaluations are more reliable but costly and noisy, as participa…