12 citations · 12 across the 1 of their papers we have counts for
1 paper
Zhihua Jin, Xingbo Wang, Furui Cheng +3
Benchmark datasets play an important role in evaluating Natural Language Understanding (NLU) models. However, shortcuts -- unwanted biases in the benchmark datasets -- can damage t…