3 papers
cs.AI2026
The Artificial Self: Characterising the landscape of AI identity
Raymond Douglas, Jan Kulveit, Ondrej Havlicek +3
Many assumptions that underpin human concepts of identity do not hold for machine minds that can be copied, edited, or simulated. We argue that there exist many different coherent…
cs.AI2026
Latent Introspection: Models Can Detect Prior Concept Injections
Theia Pearson-Vogel, Martin Vanek, Raymond Douglas +1
We uncover a latent capacity for introspection in a Qwen 32B model, demonstrating that the model can detect when concepts have been injected into its earlier context and identify w…
cs.MA2025
Multi-Agent Risks from Advanced AI
Lewis Hammond, Alan Chan, Jesse Clifton +41
The rapid development of advanced AI agents and the imminent deployment of many instances of these agents will give rise to multi-agent systems of unprecedented complexity. These s…