5 citations · 5 across the 1 of their papers we have counts for
2 papers
cs.CL2020
Robustness Testing of Language Understanding in Task-Oriented Dialog
Jiexi Liu, Ryuichi Takanobu, Jiaxin Wen +6
Most language understanding models in task-oriented dialog systems are trained on a small amount of annotated training data, and evaluated in a small set from the same distribution…
cs.CL2020★ 5 cited
CoTK: An Open-Source Toolkit for Fast Development and Fair Evaluation of Text Generation
Fei Huang, Dazhen Wan, Zhihong Shao +5
In text generation evaluation, many practical issues, such as inconsistent experimental settings and metric implementations, are often ignored but lead to unfair evaluation and unt…