1 paper
Tianpeng Zheng, Zhehan Jiang, Jiayi Liu +1
The rapid proliferation of large language models (LLMs) in healthcare creates an urgent need for scalable and psychometrically sound evaluation methods. Conventional static benchma…