6 citations · 7 across the 3 of their papers we have counts for
1 paper · 1 filter
Siqi Liu, Ian Gemp, Luke Marris +3
Evaluation has traditionally focused on ranking candidates for a specific skill. Modern generalist models, such as Large Language Models (LLMs), decidedly outpace this paradigm. Op…