works on

From the 2 of 16 linked papers with an AI index.

activity
20242026
most citedInducing language models to assert their own consciousness restores human beliefs and values

2 citations · 2 across the 6 of their papers we have counts for

collaborators
Showing cs.CYShow all

5 papers · 1 filter

cs.CY2026

Epistemic Trust as a Mechanism for Ethics Integration: Failure Modes and Design Principles from 70 Moral Imagination Workshops

Benjamin Lange, Geoff Keeling, Kyle Pedersen +4

Bottom-up responsible innovation initiatives seek to empower technology development teams to engage in ethical reflection, yet such interventions frequently fail to achieve practit…

cs.CY2025

We Need a New Ethics for a World of AI Agents

Iason Gabriel, Geoff Keeling, Arianna Manzini +1

The deployment of capable AI agents raises fresh questions about safety, human-machine relationships and social coordination. We argue for greater engagement by scientists, scholar…

cs.CY2025

Evaluating Intra-firm LLM Alignment Strategies in Business Contexts

Noah Broestl, Benjamin Lange, Cristina Voinea +2

Instruction-tuned Large Language Models (LLMs) are increasingly deployed as AI Assistants in firms for support in cognitive tasks. These AI assistants carry embedded perspectives w…

cs.CY2024

The Ethics of Advanced AI Assistants

Iason Gabriel, Arianna Manzini, Geoff Keeling +54

This paper focuses on the opportunities and the ethical and societal risks posed by advanced AI assistants. We define advanced AI assistants as artificial agents with natural langu…

cs.CY2024

A Mechanism-Based Approach to Mitigating Harms from Persuasive Generative AI

Seliem El-Sayed, Canfer Akbulut, Amanda McCroskery +17

Recent generative AI systems have demonstrated more advanced persuasive capabilities and are increasingly permeating areas of life where they can influence decision-making. Generat…