Publications (18)
Beyond Privacy Trade-offs with Structured Transparency
Andrew Trask, Emma Bluemke, Teddy Collins +5
Sociotechnical Safety Evaluation of Generative AI Systems
Laura Weidinger, Maribeth Rauh, Nahema Marchal +10
The Ethics of Advanced AI Assistants
Iason Gabriel, Arianna Manzini, Geoff Keeling +54
STAR: SocioTechnical Approach to Red Teaming Language Models
Laura Weidinger, John Mellor, Bernat Guillen Pegueroles +9
Gemini: A Family of Highly Capable Multimodal Models
Gemini Team, Rohan Anil, Sebastian Borgeaud +1340
Holistic Safety and Responsibility Evaluations of Advanced AI Models
Laura Weidinger, Joslyn Barnhart, Jenny Brennan +16
Characteristics of Harmful Text: Towards Rigorous Benchmarking of Language Models
Maribeth Rauh, John Mellor, Jonathan Uesato +9
Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Jack W. Rae, Sebastian Borgeaud, Trevor Cai +77
Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data
Nahema Marchal, Rachel Xu, Rasmi Elasmar +3
Recourse for reclamation: Chatting with generative language models
Jennifer Chien, Kevin R. McKee, Jackie Kay +1
Statistical discrimination in learning agents
Edgar A. Duéñez-Guzmán, Kevin R. McKee, Yiran Mao +9
(Unfair) Norms in Fairness Research: A Meta-Analysis
Jennifer Chien, A. Stevie Bergman, Kevin R. McKee +5
Ethical and social risks of harm from Language Models
Laura Weidinger, John Mellor, Maribeth Rauh +20
Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence
Shakir Mohamed, Marie-Therese Png, William Isaac
Toward an Evaluation Science for Generative AI Systems
Laura Weidinger, Inioluwa Deborah Raji, Hanna Wallach +7
Imagen 3
Imagen-Team-Google, :, Jason Baldridge +257
Power to the People? Opportunities and Challenges for Participatory AI
Abeba Birhane, William Isaac, Vinodkumar Prabhakaran +4
Improving alignment of dialogue agents via targeted human judgements
Amelia Glaese, Nat McAleese, Maja TrÄbacz +31