Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Reasoning Structure of Large Language Models
Frédéric Berdoz, Luca A. Lanzendörfer, Fabian Farestam +1
Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics can hide fundamentally diff…
cs.AI2024
Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies
Frédéric Berdoz, Roger Wattenhofer
While autonomous agents often surpass humans in their ability to handle vast and complex data, their potential misalignment (i.e., lack of transparency regarding their true objecti…