1 paper
Feng Wang, Yiding Sun, Jiaxin Mao +2
Large language models (LLMs) have demonstrated remarkable capabilities across various professional domains, with their performance typically evaluated through standardized benchmar…