1 paper · 1 filter
Yan Cai, Linlin Wang, Ye Wang +4
The emergence of various medical large language models (LLMs) in the medical domain has highlighted the need for unified evaluation standards, as manual evaluation of LLMs proves t…