1 paper · 1 filter
Eunsu Kim, Haneul Yoo, Guijin Son +3
As large language models (LLMs) continue to advance, the need for up-to-date and well-organized benchmarks becomes increasingly critical. However, many existing datasets are scatte…