1 paper
Qian Zhang, Panfeng Chen, Jiali Li +6
The emergence of Large Language Models (LLMs) in the medical domain has stressed a compelling need for standard datasets to evaluate their question-answering (QA) performance. Alth…