NewEvery arXiv paper, its researchers & institutions — mapped.
papers

Publications (18)

cs.CR2024

Beyond Privacy Trade-offs with Structured Transparency

Andrew Trask, Emma Bluemke, Teddy Collins +5

cs.AI2023

Sociotechnical Safety Evaluation of Generative AI Systems

Laura Weidinger, Maribeth Rauh, Nahema Marchal +10

cs.CY2024

The Ethics of Advanced AI Assistants

Iason Gabriel, Arianna Manzini, Geoff Keeling +54

cs.AI2024

STAR: SocioTechnical Approach to Red Teaming Language Models

Laura Weidinger, John Mellor, Bernat Guillen Pegueroles +9

cs.CL2025

Gemini: A Family of Highly Capable Multimodal Models

Gemini Team, Rohan Anil, Sebastian Borgeaud +1340

cs.AI2024

Holistic Safety and Responsibility Evaluations of Advanced AI Models

Laura Weidinger, Joslyn Barnhart, Jenny Brennan +16

cs.CL2022

Characteristics of Harmful Text: Towards Rigorous Benchmarking of Language Models

Maribeth Rauh, John Mellor, Jonathan Uesato +9

cs.CL2022

Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Jack W. Rae, Sebastian Borgeaud, Trevor Cai +77

cs.AI2024

Generative AI Misuse: A Taxonomy of Tactics and Insights from Real-World Data

Nahema Marchal, Rachel Xu, Rasmi Elasmar +3

cs.HC2024

Recourse for reclamation: Chatting with generative language models

Jennifer Chien, Kevin R. McKee, Jackie Kay +1

cs.LG2021

Statistical discrimination in learning agents

Edgar A. Duéñez-Guzmán, Kevin R. McKee, Yiran Mao +9

cs.CY2024

(Unfair) Norms in Fairness Research: A Meta-Analysis

Jennifer Chien, A. Stevie Bergman, Kevin R. McKee +5

cs.CL2021

Ethical and social risks of harm from Language Models

Laura Weidinger, John Mellor, Maribeth Rauh +20

cs.CY2020

Decolonial AI: Decolonial Theory as Sociotechnical Foresight in Artificial Intelligence

Shakir Mohamed, Marie-Therese Png, William Isaac

cs.AI2025

Toward an Evaluation Science for Generative AI Systems

Laura Weidinger, Inioluwa Deborah Raji, Hanna Wallach +7

cs.CV2024

Imagen 3

Imagen-Team-Google, :, Jason Baldridge +257

cs.CY2022

Power to the People? Opportunities and Challenges for Participatory AI

Abeba Birhane, William Isaac, Vinodkumar Prabhakaran +4

cs.LG2022

Improving alignment of dialogue agents via targeted human judgements

Amelia Glaese, Nat McAleese, Maja Trębacz +31