activity
20192021
most citedChallenges in Automated Debiasing for Toxic Language Detection

12 citations · 16 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CL202112 cited

Challenges in Automated Debiasing for Toxic Language Detection

Xuhui Zhou, Maarten Sap, Swabha Swayamdipta +2

Biased associations have been a challenge in the development of classifiers for detecting toxic language, hindering both fairness and accuracy. As potential solutions, we investiga…

cs.CL20204 cited

Linguistically-Informed Transformations (LIT): A Method for Automatically Generating Contrast Sets

Chuanrong Li, Lin Shengshuo, Leo Z. Liu +3

Although large-scale pretrained language models, such as BERT and RoBERTa, have achieved superhuman performance on in-distribution test sets, their performance suffers on out-of-di…

cs.CL2020

Multilevel Text Alignment with Cross-Document Attention

Xuhui Zhou, Nikolaos Pappas, Noah A. Smith

Text alignment finds application in tasks such as citation recommendation and plagiarism detection. Existing alignment methods operate at a single, predefined level and cannot lear…

cs.CL2020

RPD: A Distance Function Between Word Embeddings

Xuhui Zhou, Zaixiang Zheng, Shujian Huang

It is well-understood that different algorithms, training processes, and corpora produce different word embeddings. However, less is known about the relation between different embe…

cs.CL2019

Evaluating Commonsense in Pre-trained Language Models

Xuhui Zhou, Yue Zhang, Leyang Cui +1

Contextualized representations trained over large raw text data have given remarkable improvements for NLP tasks including question answering and reading comprehension. There have…