1 paper
Yuzhen Huang, Yuzhuo Bai, Zhihao Zhu +10
New NLP benchmarks are urgently needed to align with the rapid development of large language models (LLMs). We present C-Eval, the first comprehensive Chinese evaluation suite desi…