1 paper
Hiromi Wakaki, Yuki Mitsufuji, Yoshinori Maeda +5
We propose a new benchmark, ComperDial, which facilitates the training and evaluation of evaluation metrics for open-domain dialogue systems. ComperDial consists of human-scored re…