13 citations · 13 across the 2 of their papers we have counts for
2 papers
cs.AI2024★ 13 cited
Evaluating the Evaluator: Measuring LLMs' Adherence to Task Evaluation Instructions
Bhuvanashree Murugadoss, Christian Poelitz, Ian Drosos +5
LLMs-as-a-judge is a recently popularized method which replaces human judgements in task evaluation (Zheng et al. 2024) with automatic evaluation using LLMs. Due to widespread use…
cs.PL2024
Solving Data-centric Tasks using Large Language Models
Shraddha Barke, Christian Poelitz, Carina Suzana Negreanu +10
Large language models (LLMs) are rapidly replacing help forums like StackOverflow, and are especially helpful for non-professional programmers and end users. These users are often…