1 paper · 1 filter
Tu Vu, Kalpesh Krishna, Salaheddin Alzubi +3
As large language models (LLMs) advance, it becomes more challenging to reliably evaluate their output due to the high costs of human evaluation. To make progress towards better LL…