4 papers
LLMs in social services: How does chatbot accuracy affect human accuracy?
Jennah Gosciak, Eric Giannella, Zhaowen Guo +2
Social service programs like the Supplemental Nutrition Assistance Program (SNAP, or food stamps) have eligibility rules that can be challenging to understand. For nonprofit casewo…
Introducing AI to an Online Petition Platform Changed Outputs but not Outcomes
Isabel Corpus, Eric Gilbert, Allison Koenecke +1
The rapid integration of AI writing tools into online platforms raises critical questions about their impact on content production and outcomes. We leverage a unique natural experi…
Operationalizing Pluralistic Values in Large Language Model Alignment Reveals Trade-offs in Safety, Inclusivity, and Model Behavior
Dalia Ali, Dora Zhao, Allison Koenecke +1
Although large language models (LLMs) are increasingly trained using human feedback for safety and alignment with human values, alignment decisions often overlook human social dive…
SPHERE: Unveiling Spatial Blind Spots in Vision-Language Models Through Hierarchical Evaluation
Wenyu Zhang, Wei En Ng, Lixin Ma +5
Current vision-language models may grasp basic spatial cues and simple directions (e.g. left, right, front, back), but struggle with the multi-dimensional spatial reasoning necessa…