Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Are Bias Evaluation Methods Biased ?
Lina Berrayana, Sean Rooney, Luis Garcés-Erice +1
The creation of benchmarks to evaluate the safety of Large Language Models is one of the key activities within the trusted AI community. These benchmarks allow models to be compare…
cs.AI2025
Usage Governance Advisor: From Intent to AI Governance
Elizabeth M. Daly, Sean Rooney, Seshu Tirupathi +9
Evaluating the safety of AI Systems is a pressing concern for organizations deploying them. In addition to the societal damage done by the lack of fairness of those systems, deploy…