1 paper
Wanying Wang, Zeyu Ma, Xuhong Wang +3
As Large Language Models (LLMs) are increasingly deployed in highly specialized vertical domains, the evaluation of their domain-specific performance becomes critical. However, exi…