1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Xiang Li, Jiayi Xin, Qi Long +1
Accurate evaluation of large language models (LLMs) is crucial for understanding their capabilities and guiding their development. However, current evaluations often inconsistently…