most citedQueer In AI: A Case Study in Community-Led Participatory AI

65 citations · 117 across the 8 of their papers we have counts for

collaborators

8 papers

cs.CL20232 cited

Beyond Denouncing Hate: Strategies for Countering Implied Biases and Stereotypes in Language

Jimin Mun, Emily Allaway, Akhila Yerukola +3

Counterspeech, i.e., responses to counteract potential harms of hateful speech, has become an increasingly popular solution to address online hate speech without censorship. Howeve…

cs.CL20231 cited

FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Hyunwoo Kim, Melanie Sclar, Xuhui Zhou +4

Theory of mind (ToM) evaluations currently focus on testing models using passive narratives that inherently lack interactivity. We introduce FANToM, a new benchmark designed to str…

cs.CL20231 cited

COBRA Frames: Contextual Reasoning about Effects and Harms of Offensive Statements

Xuhui Zhou, Hao Zhu, Akhila Yerukola +4

Warning: This paper contains content that may be offensive or upsetting. Understanding the harms and offensiveness of statements requires reasoning about the social and situational…

cs.CL20234 cited

NLPositionality: Characterizing Design Biases of Datasets and Models

Sebastin Santy, Jenny T. Liang, Ronan Le Bras +2

Design biases in NLP systems, such as performance differences for different populations, often stem from their creator's positionality, i.e., views and lived experiences shaped by…

cs.CL202337 cited

Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models

Natalie Shapira, Mosh Levy, Seyed Hossein Alavi +5

The escalating debate on AI's capabilities warrants developing reliable metrics to assess machine "intelligence". Recently, many anecdotal examples were used to suggest that newer…

cs.CL2023

BiasX: "Thinking Slow" in Toxic Content Moderation with Explanations of Implied Social Biases

Yiming Zhang, Sravani Nanduri, Liwei Jiang +2

Toxicity annotators and content moderators often default to mental shortcuts when making decisions. This can lead to subtle toxicity being missed, and seemingly toxic but harmless…