1 paper
Anish Laddha, Nitesh Pradhan, Gaurav Srivastava
Large language models (LLMs) are widely used as judges for evaluating model outputs, but their high cost, latency, and opacity limit scalability. We introduce SLMJury, a framework…