1 paper · 1 filter
Adam Dejl, Jonathan Pearson
Robust and comprehensive evaluation of large language models (LLMs) is essential for identifying effective LLM system configurations and mitigating risks associated with deploying…