Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
KERAIA: An Adaptive and Explainable Framework for Dynamic Knowledge Representation and Reasoning
Stephen Richard Varey, Alessandro Di Stefano, The Anh Han
In this paper, we introduce KERAIA, a novel framework and software platform for symbolic knowledge engineering designed to address the persistent challenges of representing, reason…
cs.AI2025
Do LLMs trust AI regulation? Emerging behaviour of game-theoretic LLM agents
Alessio Buscemi, Daniele Proverbio, Paolo Bova +15
There is general agreement that fostering trust and cooperation within the AI development ecosystem is essential to promote the adoption of trustworthy AI systems. By embedding Lar…
cs.AI2024
Quantifying detection rates for dangerous capabilities: a theoretical model of dangerous capability evaluations
Paolo Bova, Alessandro Di Stefano, The Anh Han
We present a quantitative model for tracking dangerous AI capabilities over time. Our goal is to help the policy and research community visualise how dangerous capability testing c…