Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Take Goodhart Seriously: Principled Limit on General-Purpose AI Optimization
Antoine Maier, Aude Maier, Tom David
A common but rarely examined assumption in machine learning is that training yields models that actually satisfy their specified objective function. We call this the Objective Sati…
cs.AI2025
LLM Robustness Leaderboard v1 --Technical report
Pierre Peigné - Lefebvre, Quentin Feuillade-Montixi, Tom David +1
This technical report accompanies the LLM robustness leaderboard published by PRISM Eval for the Paris AI Action Summit. We introduce PRISM Eval Behavior Elicitation Tool (BET), an…