3 papers
cs.LG2026
How's it going? Reinforcement learning in language models recruits a functional welfare axis
Andy Q Han, David J. Chalmers, Pavel Izmailov
How does reinforcement learning shape a language model's internal representations? We present evidence that RL recruits a pre-existing representation of functional welfare: an esti…
cs.CY2025
Subjective Experience in AI Systems: What Do AI Researchers and the Public Believe?
Noemi Dreksler, Lucius Caviola, David Chalmers +6
We surveyed 582 AI researchers who have published in leading AI venues and 838 nationally representative US participants about their views on the potential development of AI system…
cs.CY2024
Taking AI Welfare Seriously
Robert Long, Jeff Sebo, Patrick Butlin +7
In this report, we argue that there is a realistic possibility that some AI systems will be conscious and/or robustly agentic in the near future. That means that the prospect of AI…