collaborators

6 papers

cs.CY2026

Regulating AI Agents

Kathrin Gardhouse, Amin Oueslati, Noam Kolt

AI agents -- systems that can independently take actions to pursue complex goals with only limited human oversight -- have entered the mainstream. These systems are now being widel…

cs.CY2026

Beware of GeeksBearing Gifts: Building True EU Frontier AI Sovereignty

Nick Moës, Toni Lorente, Amin Oueslati +3

Frontier artificial intelligence is reshaping all aspects of society, from economic output or military capability to democratic institutions. The EU is entering this transformation…

cs.LG2026

Open Problems in Frontier AI Risk Management

Marta Ziosi, Miro Plueckebaum, Stephen Casper +26

Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consensus resulting from the rapid pa…

cs.CY2026

Relational Archetypes: A Comparative Analysis of AV-Human and Agent-Human Interactions

Antoni Lorente, Amin Oueslati, Robin Staes-Polet

Over the last couple of years, AI Agents have gained significant traction due to substantial progress in the capabilities of underlying General Purpose AI (GPAI) models, enhanced s…

cs.CY2025

Who Should Run Advanced AI Evaluations -- AISIs?

Merlin Stein, Milan Gandhi, Theresa Kriecherbauer +2

Artificial Intelligence (AI) Safety Institutes and governments worldwide are deciding whether they evaluate advanced AI themselves, support a private evaluation ecosystem or do bot…

cs.HC2025

Lost in Moderation: How Commercial Content Moderation APIs Over- and Under-Moderate Group-Targeted Hate Speech and Linguistic Variations

David Hartmann, Amin Oueslati, Dimitri Staufer +3

Commercial content moderation APIs are marketed as scalable solutions to combat online hate speech. However, the reliance on these APIs risks both silencing legitimate speech, call…