6 citations · 6 across the 3 of their papers we have counts for
3 papers
Gaming the Answer Matcher: Examining the Impact of Text Manipulation on Automated Judgment
Manas Khatore, Sumana Sridharan, Kevork Sulahian +2
Automated answer matching, which leverages LLMs to evaluate free-text responses by comparing them to a reference answer, shows substantial promise as a scalable and aligned alterna…
Improving performance in multi-objective decision-making in Bottles environments with soft maximin approaches
Benjamin J Smith, Robert Klassert, Roland Pihlakas
Balancing multiple competing and conflicting objectives is an essential task for any artificial intelligence tasked with satisfying human values or preferences. Conflict arises bot…
Scalar reward is not enough: A response to Silver, Singh, Precup and Sutton (2021)
Peter Vamplew, Benjamin J. Smith, Johan Kallstrom +9
The recent paper `"Reward is Enough" by Silver, Singh, Precup and Sutton posits that the concept of reward maximisation is sufficient to underpin all intelligence, both natural and…