3 citations · 3 across the 1 of their papers we have counts for
1 paper · 1 filter
Aparna Elangovan, Jiayuan He, Yuan Li +1
The NLP community typically relies on performance of a model on a held-out test set to assess generalization. Performance drops observed in datasets outside of official test sets a…