3 papers
cs.LG2021
Reliability Testing for Natural Language Processing Systems
Samson Tan, Shafiq Joty, Kathy Baxter +3
Questions of fairness, robustness, and transparency are paramount to address before deploying NLP systems. Central to these concerns is the question of reliability: Can NLP systems…
cs.CL2021
Code-Mixing on Sesame Street: Dawn of the Adversarial Polyglots
Samson Tan, Shafiq Joty
Multilingual models have demonstrated impressive cross-lingual transfer performance. However, test sets like XNLI are monolingual at the example level. In multilingual communities,…
cs.CL2021
Robustness Gym: Unifying the NLP Evaluation Landscape
Karan Goel, Nazneen Rajani, Jesse Vig +6
Despite impressive performance on standard benchmarks, deep neural networks are often brittle when deployed in real-world systems. Consequently, recent research has focused on test…