Showing 2024Show all
2 papers · 1 filter
cs.CY2024
Creating a Cooperative AI Policymaking Platform through Open Source Collaboration
Aiden Lewington, Alekhya Vittalam, Anshumaan Singh +48
Advances in artificial intelligence (AI) present significant risks and opportunities, requiring improved governance to mitigate societal harms and promote equitable benefits. Curre…
cs.CL2024
Incorporating Human Explanations for Robust Hate Speech Detection
Jennifer L. Chen, Faisal Ladhak, Daniel Li +1
Given the black-box nature and complexity of large transformer language models (LM), concerns about generalizability and robustness present ethical implications for domains such as…