1 paper · 1 filter
Yuzhen Huang, Yuzhuo Bai, Zhihao Zhu +10
New NLP benchmarks are urgently needed to align with the rapid development of large language models (LLMs). We present C-Eval, the first comprehensive Chinese evaluation suite desi…