132 citations · 132 across the 4 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2023
Modeling subjectivity (by Mimicking Annotator Annotation) in toxic comment identification across diverse communities
Senjuti Dutta, Sid Mittal, Sherol Chen +6
The prevalence and impact of toxic discussions online have made content moderation crucial.Automated systems can play a vital role in identifying toxicity, and reducing the relianc…
cs.AI2023
Leveraging Contextual Counterfactuals Toward Belief Calibration
Qiuyi, Zhang, Michael S. Lee +1
Beliefs and values are increasingly being incorporated into our AI systems through alignment processes, such as carefully curating data collection principles or regularizing the lo…