219 citations · 225 across the 3 of their papers we have counts for
3 papers
cs.CL2024★ 1 cited
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
Angana Borah, Rada Mihalcea
As Large Language Models (LLMs) continue to evolve, they are increasingly being employed in numerous studies to simulate societies and execute diverse social tasks. However, LLMs a…
cs.CY2020★ 219 cited
Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims
Miles Brundage, Shahar Avin, Jasmine Wang +56
With the recent wave of progress in artificial intelligence (AI) has come a growing awareness of the large-scale impacts of AI systems, and recognition that existing regulations an…
cs.CY2020★ 5 cited
The Windfall Clause: Distributing the Benefits of AI for the Common Good
Cullen O'Keefe, Peter Cihon, Ben Garfinkel +3
As the transformative potential of AI has become increasingly salient as a matter of public and political interest, there has been growing discussion about the need to ensure that…