1 paper
Ankit Yadav, Himanshu Beniwal, Mayank Singh
Driven by the surge in code generation using large language models (LLMs), numerous benchmarks have emerged to evaluate these LLMs capabilities. We conducted a large-scale human ev…