1 paper · 1 filter
Varun Gumma, Ananditha Raghunath, Mohit Jain +1
Assessing the capabilities and limitations of large language models (LLMs) has garnered significant interest, yet the evaluation of multiple models in real-world scenarios remains…