Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Learning to Trust: Bayesian Adaptation to Varying Suggester Reliability in Sequential Decision Making
Dylan M. Asmar, Mykel J. Kochenderfer
Autonomous agents operating in sequential decision-making tasks under uncertainty can benefit from external action suggestions, which provide valuable guidance but inherently vary…
cs.AI2024
More than Marketing? On the Information Value of AI Benchmarks for Practitioners
Amelia Hardy, Anka Reuel, Kiana Jafari Meimandi +6
Public AI benchmark results are widely broadcast by model developers as indicators of model quality within a growing and competitive market. However, these advertised scores do not…