3 papers
cs.AI2025
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
Alessio Benavoli, Alessandro Facchini, Marco Zaffalon
How can we ensure that AI systems are aligned with human values and remain safe? We can study this problem through the frameworks of the AI assistance and the AI shutdown games. Th…
quant-ph2025
Connecting classical finite exchangeability to quantum theory
Alessio Benavoli, Alessandro Facchini, Marco Zaffalon
Exchangeability is a fundamental concept in probability theory and statistics. It allows to model situations where the order of observations does not matter. The classical de Finet…
cs.LG2025
The AI off-switch problem as a signalling game: bounded rationality and incomparability
Alessio Benavoli, Alessandro Facchini, Marco Zaffalon
The off-switch problem is a critical challenge in AI control: if an AI system resists being switched off, it poses a significant risk. In this paper, we model the off-switch proble…